Skip to search formSkip to main contentSkip to account menu

Apache Hadoop

Known as: HDFS, Hadoop YARN, Hadoop Distributed Filesystem 
Apache Hadoop (pronunciation: /həˈduːp/) is an open-source software framework for distributed storage and distributed processing of very large data… 
Wikipedia (opens in a new tab)

Papers overview

Semantic Scholar uses AI to extract papers important to this topic.
2017
2017
Hadoop-MapReduce is a popular distributed computing model that has been deployed on large clusters like those owned by Yahoo and… 
2017
2017
Hadoop distributed file system (HDFS) and MapReduce model have become popular technologies for large‐scale data organization and… 
2017
2017
Data analytics is undergoing a revolution in many scientific domains, and demands cost-effective parallel data analysis… 
2014
2014
Big Data technology is becoming ubiquitous and depicting key attention of researchers in almost all areas. VANET is a special… 
2014
2014
The emergence of cloud computing solves the problems that traditional data mining algorithms encounter when dealing with large… 
2013
2013
We currently live in the data age. It’s not easy to measure the total volume of structured and unstructured data that require… 
2012
2012
The programming paradigm Map-Reduce and its main open-source implementation, Hadoop, have had an enormous impact on large scale… 
2011
2011
Query processing using mostly various NoSQL languages becomes a significant application area for Hadoop. Despite significant work… 
2009
2009
With the development of cloud computing, Data security becomes more and more important in cloud computing. This paper analyses…