|
所在平台: Coursera |
课程主页: https://www.coursera.org/learn/data-lakes-data-warehouses-gcp
课程评论:没有评论
课程名称:使用 Google Cloud 现代化数据湖和数据仓库 概述:数据管道的两个关键组成部分是数据湖和数据仓库。本课程重点介绍每种存储类型的使用案例,并深入探讨 Google Cloud 上可用的数据湖和数据仓库解决方案的技术细节。同时,本课程阐述了数据工程师的角色、成功的数据管道对业务运营的好处,并探讨了为何数据工程应该在云环境中进行。这是 Google Cloud 数据工程系列的第一门课程。完成本课程后,您可以报名参加“在 Google Cloud 上构建批量数据管道”课程。 课程大纲: 1. 介绍 - 描述:本模块介绍了 Google Cloud 数据工程系列和《使用 Google Cloud 现代化数据湖和数据仓库》课程。 2. 数据工程简介 - 描述:本模块讨论了数据工程的角色,并阐明了为什么数据工程应该在云中进行的原因。 3. 构建数据湖 - 描述:在本模块中,我们描述了数据湖的概念,并介绍了如何使用 Google Cloud 的 Cloud Storage 作为数据湖。 4. 构建数据仓库 - 描述:本模块讨论了 BigQuery 作为 Google Cloud 上的数据仓库选项。 5. 总结 - 描述:总结关键学习要点。
Name:Introduction
Description:This module introduces the Data Engineering on Google Cloud source series and this Modernizing Data Lakes and Data Warehouses with Google Cloud course.
Name:Introduction to Data Engineering
Description:This module discusses the role of data engineering and motivates the claim why data engineering should be done in the Cloud
Name:Building a Data Lake
Description:In this module, we describe what data lake is and how to use Cloud Storage as your data lake on Google Cloud.
Name:Building a data warehouse
Description:In this module, we talk about BigQuery as a data warehousing option on Google Cloud
Name:Summary
Description:A summary of the key learning points
The two key components of any data pipeline are data lakes and warehouses. This course highlights use-cases for each type of storage and dives into the available data lake and warehouse solutions on Google Cloud in technical detail. Also, this course describes the role of a data engineer, the benefits of a successful data pipeline to business operations, and examines why data engineering should be done in a cloud environment. This is the first course of the Data Engineering on Google Cloud series. After completing this course, enroll in the Building Batch Data Pipelines on Google Cloud course.