开始时间: 02/16/2019 持续时间: Unknown
This specialization is made for people working with data (either small or big). If you are a Data Analyst, Data Scientist, Data Engineer or Data Architect (or you want to become one) — don’t miss the opportunity to expand your knowledge and skills in the field of data engineering and data analysis on the large scale. In four concise courses you will learn the basics of Hadoop, MapReduce, Spark, methods of offline data processing for warehousing, real-time data processing and large-scale machine learning. And Capstone project for you to build and deploy your own Big Data Service (make your portfolio even more competitive). Over the course of the specialization, you will complete progressively harder programming assignments (mostly in Python). Make sure, you have some experience in it. This course will master your skills in designing solutions for common Big Data tasks: - creating batch and real-time data processing pipelines, - doing machine learning at scale, - deploying machine learning models into a production environment — and much more! Join some of best hands-on big data professionals, who know, their job inside-out, to learn the basics, as well as some tricks of the trade, from them. Special thanks to Prof. Mikhail Roytberg (APT dept., MIPT), Oleg Sukhoroslov (PhD, Senior Researcher, IITP RAS), Oleg Ivchenko (APT dept., MIPT), Pavel Akhtyamov (APT dept., MIPT), Vladimir Kuznetsov, Asya Roitberg, Eugene Baulin, Marina Sudarikova.
Big Data Essentials: HDFS, MapReduce and Spark RDD
Big Data Analysis: Hive, Spark SQL, DataFrames and GraphFrames
Big Data Applications: Machine Learning at Scale
Big Data Applications: Real-Time Streaming
俄罗斯搜索巨头Yandex推出的面向数据工程师的大数据专项课程系列（Big Data for Data Engineers Specialization） ，这个系列包括4门子课程和1门毕业项目课程，涵盖HDFS，MapReduce, Spark, Hive, Spark SQL, 大规模机器学习，实时流处理等，感兴趣的同学可以关注: Build Your Data Engineering Skills-Learn how to tame the big data beast with the most popular tools assisted by top-notch practitioners