Master Apache Spark (Scala) for Data Engineers

所在平台: Udemy

课程主页: https://www.udemy.com/course/learn-advance-spark-beginner-to-expert-scala/

课程评论:没有评论

第一个写评论        关注课程

课程简介

Coursera 课程总结:《Master Apache Spark (Scala) for Data Engineers》 本课程旨在为数据工程师和架构师提供一个从基础到高级的 Apache Spark 3.x 学习路径,以最有效和简洁的方式掌握该技术。无论您是初学者还是已有 Spark 基础,本课程都能为您带来价值。 课程将深入探讨 Spark 内部机制、Datasets、执行计划,并涵盖在 IntelliJ IDE 中进行开发以及在 EMR 集群(AWS 云)上运行 Spark 的实践操作。课程无需任何 Apache Spark 或 Hadoop 基础知识。Spark 的架构和基本概念将得到详细解释,以帮助您全面理解课程内容。 本课程将使用 Scala 编程语言,这是处理 Apache Spark 的最佳选择。 **课程内容包括:** * 大数据生态系统入门 * Spark 内部机制详解 * 理解 Spark Driver 和 Executor * 执行计划的深度解析 * 在本地/Google Cloud 上设置开发环境 * Spark DataFrames 的使用 * IntelliJ IDE 的应用 * 在 EMR 集群(AWS Cloud)上运行 Spark * 高级 DataFrame 示例 * RDD 的使用 * RDD 示例 完成本课程后,您将能够自信地回答 Spark 面试问题,并能编写代码在几分钟内分析海量数据。

课程评论(0条)

课程详情

This course is designed in such a manner to cover basics to advanced concept to learn Apache Spark 3.x in most efficient and concise manner. This course will be beneficial for beginners as well as for those who already know Apache Spark. It covers in-depth details about spark internals, datasets, execution plan, Intellij IDE, EMR cluster with lots of hands on. This course is designed for Data Engineers and Architects who are willing to design and develop a Bigdata Engineering Projects using Apache Spark. It does not require any prior knowledge of Apache Spark or Hadoop. Spark Architecture and fundamental concepts are explained in details to help you grasp the content of this course. This course uses the Scala programming language which is the best language to work with Apache Spark. This course covers:Intro to Big data ecosystemSpark Internals in detailsUnderstanding Spark Drivers, executors.Understanding Execution plan in detailsSetting up environment on Local/Google cloudWorking with Spark DataframesWorking with Intellij IDERunning Spark on EMR cluster (AWS Cloud)Advanced Dataframe examplesWorking with RDDRDD examplesBy the end of this course, you'll be able to answer any spark interview question and will be able to run code that analyzes gigabytes worth of information in Apache Spark in a matter of minutes.

课程标签

0人关注该课程

主题相关的课程