chronicle

eine chronik in blogs


2008.11
Sorting 1PB with MapReduce Sorting 1PB with MapReduce - Google sortiert ein Petabyte.

At Google we are fanatical about organizing the world’s information. As a result, we spend a lot of time finding better ways to sort information using MapReduce, a key component of our software infrastructure that allows us to run multiple processes simultaneously. MapReduce is a perfect solution for many of the computations we run daily, due in large part to its simplicity, applicability to a wide range of real-world computing tasks, and natural translation to highly scalable distributed implementations that harness the power of thousands of computers.

# google