|
所在平台: Udemy |
课程主页: https://www.udemy.com/course/learn-apache-spark-with-python/
课程评论:没有评论
**课程名称:** 使用Python学习Apache Spark **课程概述:** Apache Spark是当今最热门的大数据技能。越来越多的组织正在采用Apache Spark来构建他们的大数据处理和分析应用程序,对Apache Spark专业人才的需求正在迅速增长。学习Apache Spark是获得好工作、提高工作质量和获得最高薪酬的绝佳途径。 您可能已经知道Apache Spark是一个快速通用的数据处理引擎,内置了流处理、SQL、机器学习和图处理模块。它以其速度、易用性、通用性以及几乎无处不在的运行能力而闻名。尽管Spark是数据工程师需求量最大的工具之一,但数据科学家在进行探索性数据分析、特征提取、监督学习和模型评估时也可以从Spark中受益。 本课程将涵盖使用Python学习Apache Spark的许多主题,包括: * Spark成为大数据和数据科学强大工具的原因 * 学习Spark基础知识,包括弹性分布式数据集(RDD)、Spark操作(Actions)和转换(Transformations) * 探索Spark SQL,支持CSV、JSON和MySQL(JDBC)数据源 * 提供方便的链接以下载所有源代码
Apache Spark is the hottest Big Data skill today. More and more organizations are adapting Apache Spark for building their big data processing and analytics applications and the demand for Apache Spark professionals is sky rocketing. Learning Apache Spark is a great vehicle to good jobs, better quality of work and the best remuneration packages. You might already know Apache Spark as a fast and general engine for big data processing, with built-in modules for streaming, SQL, machine learning and graph processing. It's well-known for its speed, ease of use, generality and the ability to run virtually everywhere. And even though Spark is one of the most asked tools for data engineers, also data scientists can benefit from Spark when doing exploratory data analysis, feature extraction, supervised learning and model evaluation. The course will cover many more topics of Apache Spark with Python including-What makes Spark a power tool of Big Data and Data Science?Learn the fundamentals of Spark including Resilient Distributed Datasets, Spark Actions and TransformationsExplore Spark SQL with CSV, JSON and mySQL (JDBC) data sourcesConvenient links to download all source code