Level · 30 daysSoon
30 Days of Big Data
When data won’t fit in memory — distributed processing with Spark and the patterns of big-data pipelines.
- Module 1Why distributed computing
- Module 2MapReduce thinking
- Module 3Spark fundamentals
- Module 4DataFrames at scale
- Module 5Partitioning & shuffles
- Module 6A big-data pipeline