|
所在平台: Udemy |
课程主页: https://www.udemy.com/course/aws-etl-glue-data-pipeline-and-athena-fundamentals/
课程评论:没有评论
课程名称:AWS ETL [Glue, Data Pipeline 和 Athena] 基础知识 课程概述: 欢迎参加本AWS数据管道课程,我们非常高兴能够向您推出这门课程。该课程将带领您从初学者成长为在云端使用AWS Data Pipeline和AWS Glue创建ETL管道的专家。通过实例学习,我们将演示所有讨论的概念,您可以亲自实践。我的逐步培训将引导您进入AWS ETL管道的世界。 Amazon提供的ETL(提取、转换和加载)工具可以与不同的数据存储集成。一旦准备好数据,AWS允许您使用Athena对其进行查询,然后使用Quicksight构建仪表板。重要的是,您需要知道什么时候使用哪个工具。EMR在与Hadoop结合使用时效果很好,而Glue则适用于快速管道的构建。 为什么要学习AWS Data Pipeline/Glue/EMR/Athena?当您需要清洗和转换分散在多个数据源中的数据以便于分析时,应该使用什么工具进行可扩展的数据集成?AWS Data Pipeline是一个网络服务,使客户能够创建自动化的数据运输和转换操作。换句话说,它提供了作为服务的数据提取、加载和转换。用户无需构建昂贵的ETL或ELT平台,而是可以利用Amazon的云环境。随着业务活动的增长,数据量也在增加,因此管道的可扩展性至关重要。 AWS Glue与Glue Studio和Glue Data Brew等其他工具结合使用,允许您构建具有不同程度自定义的管道。AWS Glue允许您构建自定义ETL管道,而Glue Studio提供了一个用于监控的一体化工具。AWS Glue Data Brew是一个新的工具,它允许您在无需编码的情况下构建管道,您可以在几次点击中应用超过250种转化! AWS Glue Data Classify提供了您数据的一致视图,使您可以适当地清洗、丰富和编目数据。这也确保您的数据可以被搜索、查询,并立即准备好进行ETL。理解何时使用AWS EMR以及何时使用AWS Glue,从工程到数据再到分析,这为企业团队提供了巨大的潜力。最后,我们将通过一些真实案例来研究AWS Glue或Data Pipeline可以如何保护我们的工作。我们确保您始终学习最佳实践。 在课程结束时,您将能够像专家一样设置和管理AWS中的数据转换。所有内容均有良好的文档和分类,您可以轻松找到所需信息。作业和测验将确保您保持进度并测试您的知识。课程将结合理论和实践示例。
Welcome to this AWS Data Pipeline CourseWe are very excited to get this course out to you. This course will take you from being a beginner to an expert in creating ETL Pipelines on the cloud using AWS Data Pipeline and AWS Glue.In this course you will Learn by example, where we demonstrate all the concepts discussed so that you can see them working, and you can try them out for yourself as well.My step-by-step training will initiate you into the world of AWS ETL Pipelines.Amazon provides ETL (Extract, Transform and Load) tools which integrate with different data storages. Once you prepare the data AWS allows you to query it using Athena and then build dashboards using Quicksight. What is important is you know which tool to use when. EMR is great when used with Hadoop but Glue is great for quick pipeline.Why Learn AWS Data Pipeline/Glue/EMR/Athena?Another Question: What tool should you use for Scalable Data Integration when you have data dispersed in multiple data sources and need it to be cleaned and transformed for analysis?AWS Data Pipeline is a web service that allows customers to create automated data transport and transformation operations. In other words, it provides data extraction, loading, and transformation as a service. To use their data, users don't need to build an expensive ETL or ELT platform; instead, they may use Amazon's cloud environment. Data keeps growing as business activity grows, so the scalability of your pipeline is important.AWS Glue coupled with other tools like Glue Studio and Glue Data Brew allows you to build pipelines with varying amount of customization. AWS Glue allows you to build custom ETL pipelines while Glue Studio provides a UI tool to monitor everything. AWS Glue Data Brew is a new tool which allows you to build the pipelines without any coding. It has more than 250 transformations you can apply with just a few clicks!AWS Glue Data Classify provides a consistent view of your data, allowing you to appropriately clean, enrich, and catalogue it. This also assures that your data is searchable, queryable, and ETL-ready right away. Understand when to use AWS EMR and when to use AWS Glue. From engineering to data to analytics, it offers immense potential for teams across corporate businesses. Finally we take a look at some real life examples to study where and how AWS Glue or Data Pipeline can help us. We ensure you learn the best practices at all times.By the end of the course you will have set up and learnt to manage data transformations in AWS like a pro.Everything is well documented and separated, so you can find what you need. Assignments and Quizzes will make sure you stay on track and test your knowledge. The course will have a combination of theory and practical examples.