課程目錄: 大數據整合與處理培訓

        4401 人關注
        (78637/99817)
        課程大綱:

        大數據整合與處理培訓

         

         

         

        1

        Welcome to Big Data Integration and Processing
        Welcome to the third course in the Big Data Specialization. This week you will be introduced to basic concepts
        in big data integration and processing. You will be guided through installing the Cloudera VM,
        downloading the data sets to be used for this course, and learning how to run the Jupyter server.

        Retrieving Big Data (Part 1)
        This module covers the various aspects of data retrieval and relational querying. You will also be introduced to the Postgres database.

        Retrieving Big Data (Part 2)
        This module covers the various aspects of data retrieval for NoSQL data, as well as data aggregation and working with data frames. You will be
        introduced to MongoDB and Aerospike, and you will learn how to use Pandas to retrieve data from them.

        Big Data Integration
        In this module you will be introduced to data integration tools including Splunk and Datameer, and you will gain some practical insight into how
        information integration processes are carried out.

        Processing Big Data
        This module introduces Learners to big data pipelines and workflows as well as processing and analysis of big data using Apache Spark.

        Big Data Analytics using Spark
        In this module, you will go deeper into big data processing by learning the inner workings of the Spark Core. You will be introduced to two key
        tools in the Spark toolkit: Spark MLlib and GraphX.

        Learn By Doing: Putting MongoDB and Spark to Work
        In this module you will get some practical hands-on experience applying what you learned about Spark and MongoDB to analyze Twitter data.