Skip to Main content Skip to Navigation
New interface
Conference papers

BlobSeer: Bringing High Throughput under Heavy Concurrency to Hadoop Map-Reduce Applications

Bogdan Nicolae 1 Diana Moise 1 Gabriel Antoniu 1, * Luc Bougé 1 Matthieu Dorier 1 
* Corresponding author
1 KerData - Scalable Storage for Clouds and Beyond
Inria Rennes – Bretagne Atlantique , IRISA-D1 - SYSTÈMES LARGE ÉCHELLE
Abstract : Hadoop is a software framework supporting the Map-Reduce programming model. It relies on the Hadoop Distributed File System (HDFS) as its primary storage system. The efficiency of HDFS is crucial for the performance of Map-Reduce applications. We substitute the original HDFS layer of Hadoop with a new, concurrency-optimized data storage layer based on the BlobSeer data management service. Thereby, the efficiency of Hadoop is significantly improved for data-intensive Map-Reduce applications, which naturally exhibit a high degree of data access concurrency. Moreover, BlobSeer's features (built-in versioning, its support for concurrent append operations) open the possibility for Hadoop to further extend its functionalities. We report on extensive experiments conducted on the Grid'5000 testbed. The results illustrate the benefits of our approach over the original HDFS-based implementation of Hadoop.
Complete list of metadata

Cited literature [9 references]  Display  Hide  Download
Contributor : Luc Bougé Connect in order to contact the contributor
Submitted on : Monday, February 15, 2010 - 5:28:20 PM
Last modification on : Wednesday, February 2, 2022 - 3:50:46 PM
Long-term archiving on: : Thursday, June 30, 2011 - 12:08:31 PM


Files produced by the author(s)



Bogdan Nicolae, Diana Moise, Gabriel Antoniu, Luc Bougé, Matthieu Dorier. BlobSeer: Bringing High Throughput under Heavy Concurrency to Hadoop Map-Reduce Applications. 24th IEEE International Parallel and Distributed Processing Symposium (IPDPS 2010), IEEE and ACM, Apr 2010, Atlanta, United States. ⟨10.1109/IPDPS.2010.5470433⟩. ⟨inria-00456801⟩



Record views


Files downloads