General Introduction to Big DataCovers data science tools, Hadoop, Spark, data lake ecosystems, CAP theorem, batch vs. stream processing, HDFS, Hive, Parquet, ORC, and MapReduce architecture.
Data Wrangling with HadoopCovers data wrangling techniques using Hadoop, focusing on row versus column-oriented databases, popular storage formats, and HBase-Hive integration.
OpenShift Survival GuideProvides a survival guide for OpenShift, covering node setup, service management, configuration handling, and issue troubleshooting.
Data Science EssentialsCovers the essentials of data science, including data handling, visualization, and analysis, emphasizing practical skills and active engagement.