Advanced Methods in Data Science and Big Data Analytics (AMDSBDA) – Outline

Outline detalhado do curso

Module 1: MapReduce and Hadoop
  • Lesson 1: The MapReduce Framework
  • Lesson 2: Apache Hadoop
  • Lesson 3: Hadoop Distributed File System
  • Lesson 4: YARN
Module 2: Hadoop Ecosystem and NoSQL
  • Lesson 1: Hadoop Ecosystem
  • Lesson 2: Pig
  • Lesson 3: Hive
  • Lesson 4: NoSQL - Not Only SQL
  • Lesson 5: HBase
  • Lesson 6: Spark
Module 3: Natural Language Processing
  • Lesson 1: Introduction tNLP
  • Lesson 2: Text Preprocessing
  • Lesson 3: TFIDF
  • Lesson 4: Beyond Bag of Words
  • Lesson 5: Language Modeling
  • Lesson 6: POS Tagging and HMM
  • Lesson 7: Sentiment Analysis and Topic Modeling
Module 4: Social Network Analysis
  • Lesson 1: Introduction tSNA and Graph Theory
  • Lesson 2: Most Important Nodes
  • Lesson 3: Communities and Small World
  • Lesson 4: Network Problems and SNA Tools
Module 5: Data Science Theory and Methods
  • Lesson 1: Simulation
  • Lesson 2: Random Forests
  • Lesson 3: Multinomial Logistic Regression
Module 6: Data Visualization
  • Lesson 1: Perception and Visualization
  • Lesson 2: Visualization of Multivariate Data Module