基于Google云平台的批量数据管道构建培训
Introduction
In this module, we introduce the course and agenda
ntroduction to Batch Data Pipelines
This module reviews different methods of data loading: EL, ELT and ETL and when to use what
Executing Spark on Cloud Dataproc
This module shows how to run Hadoop on Cloud Dataproc, how to leverage GCS,
and how to optimize your Dataproc jobs.
Manage Data Pipelines with Cloud Data Fusion and Cloud Composer
This module shows how to manage data pipelines with Cloud Data Fusion and Cloud Composer.
Serverless Data Processing with Cloud Dataflow
This module covers using Cloud Dataflow to build your data processing pipelines
Summary
This module reviews the topics covered in this course