|
所在平台: Coursera专项课程 |
课程主页: https://www.coursera.org/specializations/nosql-big-data-and-spark-foundations
课程评论:没有评论
课程名称:NoSQL、大数据与Spark基础 课程概述: 本课程旨在教你如何使用NoSQL数据库进行数据插入、更新、删除、查询、索引、聚合以及分片/分区。您将获得MongoDB、Apache Cassandra和IBM Cloudant等NoSQL数据库的实际操作经验。同时,您也将建立对大数据的基础知识,并通过Apache Hadoop、MapReduce、Apache Spark、Spark SQL和Kubernetes的实验室实践来加深理解。您将学习如何进行数据的提取、转换和加载(ETL)处理,以及如何使用Apache Spark进行机器学习模型的训练和部署。 您将掌握的技能: - Apache Hadoop - MongoDB - 大数据 - Apache Spark - NoSQL数据库 - Cloudant - Cassandra - Spark SQL - Spark ML 专业介绍: 在数据管理行业,拥有NoSQL技能的大数据工程师和专业人士非常抢手。本专门化课程旨在帮助学生建立使用大数据、Apache Spark和NoSQL数据库的基本技能。课程包括三个内容丰富的部分,涵盖流行的NoSQL数据库,如MongoDB和Apache Cassandra,广泛使用的Apache Hadoop大数据工具生态系统,以及Apache Spark大规模数据处理的分析引擎。 学习内容包括: - 各类NoSQL数据存储的概述,以及使用IBM Cloudant、MongoDB和Cassandra等数据库的实操。 - 数据管理任务:创建和复制数据库、插入、更新、删除、查询、索引、聚合和分片数据。 - 大数据技术基础知识,如Hadoop、MapReduce、HDFS、Hive和HBase,以及深入了解Apache Spark、Spark数据框架、Spark SQL、PySpark、Spark应用程序界面及使用Kubernetes扩展Spark的能力。 - 使用Spark进行ETL处理和机器学习任务的实践。 该专门化适合希望入门NoSQL和大数据领域的初学者,不论是数据工程师、软件开发人员、IT架构师、数据科学家或IT经理。 应用学习项目: 本专门化着重于实践学习,每门课程都包含动手实验,以便学员在讲座中学习到的NoSQL和大数据技能进行实践应用。在第一门课程中,您将与多个NoSQL数据库(如MongoDB、Apache Cassandra和IBM Cloudant)进行实际操作。接着,您将启动使用Docker构建的Hadoop集群并运行MapReduce作业,用Jupyter Notebook在Python环境中操作Spark,构建Spark技能,并在最后一门课程中使用IBM Watson进行ETL处理和机器学习模型的训练与部署。 证书: 完成课程后可获得可分享证书。所有课程均为在线课程,您可立即开始学习并按自己的节奏进行安排,设定灵活的截止日期。该课程适合已具备基本计算机和数据素养技能,并有一定Python和SQL编程基础的学习者。 预计完成时间: 大约需时4个月,每周建议学习3小时。 Available languages: 英语,多种语言字幕选项。
Course Link: https://www.coursera.org/learn/introduction-to-nosql-databases
Name:Introduction to NoSQL Databases
Description:Offered by IBM. Get started with NoSQL Databases with this beginner-friendly introductory course! This course will provide technical, ... Enroll for free.
Course Link: https://www.coursera.org/learn/introduction-to-big-data-with-spark-hadoop
Name:Introduction to Big Data with Spark and Hadoop
Description:Offered by IBM. This self-paced IBM course will teach you all about big data! You will become familiar with the characteristics of big data ... Enroll for free.
Course Link: https://www.coursera.org/learn/machine-learning-with-apache-spark
Name:Machine Learning with Apache Spark
Description:Offered by IBM. Explore the exciting world of machine learning with this IBM course. Start by learning ML fundamentals before unlocking ... Enroll for free.
What you will learn
Work with NoSQL databases to insert, update, delete, query, index, aggregate, and shard/partition data.
Develop hands-on NoSQL experience working with MongoDB, Apache Cassandra, and IBM Cloudant.
Develop foundational knowledge of Big Data and gain hands-on lab experience using Apache Hadoop, MapReduce, Apache Spark, Spark SQL, and Kubernetes.
Perform Extract, Transform and Load (ETL) processing and Machine Learning model training and deployment with Apache Spark.
Skills you will gain
Apache Hadoop
Mongodb
Big Data
Apache Spark
NoSQL Databases
NoSQL
Cloud Database
Cloudant
Cassandra
SparkSQL
SparkML
About this Specialization
7,226
recent views
Big Data Engineers and professionals with NoSQL skills are highly sought after in the data management industry. This Specialization is designed for those seeking to develop fundamental skills for working with Big Data, Apache Spark, and NoSQL databases. Three information-packed courses cover popular NoSQL databases like MongoDB and Apache Cassandra, the widely used Apache Hadoop ecosystem of Big Data tools, as well as Apache Spark analytics engine for large-scale data processing.
You start with an overview of various categories of NoSQL (Not only SQL) data repositories, and then work hands-on with several of them including IBM Cloudant, MonogoDB and Cassandra. You’ll perform various data management tasks, such as creating & replicating databases, inserting, updating, deleting, querying, indexing, aggregating & sharding data. Next, you’ll gain fundamental knowledge of Big Data technologies such as Hadoop, MapReduce, HDFS, Hive, and HBase, followed by a more in depth working knowledge of Apache Spark, Spark Dataframes, Spark SQL, PySpark, the Spark Application UI, and scaling Spark with Kubernetes. In the final course, you will learn to work with Spark Structured Streaming Spark ML - for performing Extract, Transform and Load processing (ETL) and machine learning tasks.
This specialization is suitable for beginners in the fields of NoSQL and Big Data – whether you are or preparing to be a Data Engineer, Software Developer, IT Architect, Data Scientist, or IT Manager.
Applied Learning Project
The emphasis in this specialization is on learning by doing. As such, each course includes hands-on labs to practice & apply the NoSQL and Big Data skills you learn during lectures.
In the first course, you will work hands-on with several NoSQL databases- MongoDB, Apache Cassandra, and IBM Cloudant to perform a variety of tasks: creating the database, adding documents, querying data, utilizing the HTTP API, performing Create, Read, Update & Delete (CRUD) operations, limiting & sorting records, indexing, aggregation, replication, using CQL shell, keyspace operations, & other table operations.
In the next course, you’ll launch a Hadoop cluster using Docker and run Map Reduce jobs. You’ll
explore working with Spark using Jupyter notebooks on a Python kernel. You’ll build your Spark skills using DataFrames, Spark SQL, and scale your jobs using Kubernetes.
In the final course you will use Spark for ETL processing, and Machine Learning model training and deployment using IBM Watson.
Shareable Certificate
Shareable Certificate
Earn a Certificate upon completion
100% online courses
100% online courses
Start instantly and learn at your own schedule.
Flexible Schedule
Flexible Schedule
Set and maintain flexible deadlines.
Beginner Level
Beginner Level
The courses in the specialization require that you have basic computer and data literacy skills, as well as some programming background with languages such as with Python and SQL. No prior knowledge or experience of Big Data and NoSQL is required.
Hours to complete
Approximately 4 months to complete
Suggested pace of 3 hours/week
Available languages
English
Subtitles: English
Shareable Certificate
Shareable Certificate
Earn a Certificate upon completion
100% online courses
100% online courses
Start instantly and learn at your own schedule.
Flexible Schedule
Flexible Schedule
Set and maintain flexible deadlines.
Beginner Level
Beginner Level
The courses in the specialization require that you have basic computer and data literacy skills, as well as some programming background with languages such as with Python and SQL. No prior knowledge or experience of Big Data and NoSQL is required.
Hours to complete
Approximately 4 months to complete
Suggested pace of 3 hours/week
Available languages
English
Subtitles: English