课程目录: 基于Google云平台的批量数据管道构建培训

4401 人关注
(78637/99817)
课程大纲:

基于Google云平台的批量数据管道构建培训

 

 

 

Introduction

In this module, we introduce the course and agenda

ntroduction to Batch Data Pipelines

This module reviews different methods of data loading: EL, ELT and ETL and when to use what

Executing Spark on Cloud Dataproc

This module shows how to run Hadoop on Cloud Dataproc, how to leverage GCS,

and how to optimize your Dataproc jobs.

 

Manage Data Pipelines with Cloud Data Fusion and Cloud Composer

This module shows how to manage data pipelines with Cloud Data Fusion and Cloud Composer.

 

Serverless Data Processing with Cloud Dataflow

This module covers using Cloud Dataflow to build your data processing pipelines

Summary

This module reviews the topics covered in this course