Big Data Integration and Processing

Big Data Integration and Processing

Archived
课程
en
英语
20 时
此内容评级为 0/5

你无法访问存档的 讲座

来源
  • 来自www.coursera.org
状况
  • 自定进度
  • 免费获取
  • 收费证书
更多信息
  • 6 序列
  • 等级 介绍

你无法访问存档的 讲座

Their employees are learning daily with Edflex

  • Safran
  • Air France
  • TotalEnergies
  • Generali
Learn more

课程详情

教学大纲

  • Week 1 - Welcome to Big Data Integration and Processing
    Welcome to the third course in the Big Data Specialization. This week you will be introduced to basic concepts in big data integration and processing. You will be guided through installing the Cloudera VM, downloading the data sets to be used for this course, ...
  • Week 1 - Retrieving Big Data (Part 1)
    This module covers the various aspects of data retrieval and relational querying. You will also be introduced to the Postgres database.
  • Week 2 - Retrieving Big Data (Part 2)
    This module covers the various aspects of data retrieval for NoSQL data, as well as data aggregation and working with data frames. You will be introduced to MongoDB and Aerospike, and you will learn how to use Pandas to retrieve data from them.
  • Week 3 - Big Data Integration
    In this module you will be introduced to data integration tools including Splunk and Datameer, and you will gain some practical insight into how information integration processes are carried out.
  • Week 4 - Processing Big Data
    This module introduces Learners to big data pipelines and workflows as well as processing and analysis of big data using Apache Spark.
  • Week 5 - Big Data Analytics using Spark
    In this module, you will go deeper into big data processing by learning the inner workings of the Spark Core. You will be introduced to two key tools in the Spark toolkit: Spark MLlib and GraphX.
  • Week 6 - Learn By Doing: Putting MongoDB and Spark to Work
    In this module you will get some practical hands-on experience applying what you learned about Spark and MongoDB to analyze Twitter data.

先决条件

没有。

讲师

Ilkay Altintas
Chief Data Science Officer
San Diego Supercomputer Center

Amarnath Gupta
Director, Advanced Query Processing Lab
San Diego Supercomputer Center (SDSC)

编辑

加州大学圣地亚哥分校是位于加利福尼亚州圣地亚哥的一所公立赠地研究型大学。加州大学圣地亚哥分校成立于 1960 年,位于斯克里普斯海洋学研究所附近,是加州大学十个校区中最南端的一个,提供 200 多个本科和研究生学位课程,在校本科生 33,096 人,研究生 9,872 人。

加州大学圣地亚哥分校被认为是世界上最好的大学之一。多份出版物将加州大学圣地亚哥分校的生物科学系和计算机科学系评为世界前十名。

平台

Coursera是一家数字公司,提供由位于加利福尼亚州山景城的计算机教师Andrew Ng和达芙妮科勒斯坦福大学创建的大型开放式在线课程。

Coursera与顶尖大学和组织合作,在线提供一些课程,并提供许多科目的课程,包括:物理,工程,人文,医学,生物学,社会科学,数学,商业,计算机科学,数字营销,数据科学 和其他科目。

此内容评级为 4.5/5
(没有评论)
此内容评级为 4.5/5
(没有评论)
完成这个资源,写一篇评论