Learn More
To improve data availability and resilience MapReduce frameworks use file systems that replicate data <i>uniformly</i>. However, analysis of job logs from a large production cluster shows wide disparity in data popularity. Machines and racks storing popular content become bottlenecks; thereby increasing the completion times of jobs accessing this data even(More)
  • 1