Loading timeline…
20042000s
MapReduce / Big Data
A programming model for processing massive datasets across distributed clusters.
Why It Was Important
Published by Google engineers Jeffrey Dean and Sanjay Ghemawat, MapReduce allowed developers to write simple code to process petabytes of data across thousands of commodity servers without worrying about network failure or fault tolerance. It unleashed the 'Big Data' era, enabling the collection of training data required for modern AI.
Who Invented It
Jeffrey Dean & Sanjay Ghemawat
Legendary distributed systems engineers at Google.
Applications
- Search Engine Indexing
- Log Analysis
- Hadoop Ecosystem
Key Papers
- MapReduce: Simplified Data Processing on Large Clusters
Jeffrey Dean, Sanjay Ghemawat · Communications of the ACM · 2008
Videos
MapReduce - Computerphile
Computerphile