Skip to search form
Skip to main content
Skip to account menu
Semantic Scholar
Semantic Scholar's Logo
Search 238,332,909 papers from all fields of science
Search
Sign In
Create Free Account
Apache Hadoop
Known as:
HDFS
, Hadoop YARN
, Hadoop Distributed Filesystem
Expand
Apache Hadoop (pronunciation: /həˈduːp/) is an open-source software framework for distributed storage and distributed processing of very large data…
Expand
Wikipedia
(opens in a new tab)
Create Alert
Alert
Related topics
Related topics
50 relations
Amazon Elastic Compute Cloud (EC2)
Apache Flume
Apache Giraph
Apache Gora
Expand
Papers overview
Semantic Scholar uses AI to extract papers important to this topic.
2017
2017
Towards a Resource Aware Scheduler in Hadoop
M. Yong
,
Nitin Garegrat
,
Shiwali Mohan
2017
Corpus ID: 7782658
Hadoop-MapReduce is a popular distributed computing model that has been deployed on large clusters like those owned by Yahoo and…
Expand
2017
2017
XHAMI – extended HDFS and MapReduce interface for Big Data image processing applications in cloud computing environments
Raghavendra Kune
,
P. Konugurthi
,
Arun Agarwal
,
Raghavendra Rao Chillarige
,
R. Buyya
Software, Practice & Experience
2017
Corpus ID: 17018548
Hadoop distributed file system (HDFS) and MapReduce model have become popular technologies for large‐scale data organization and…
Expand
2017
2017
A Data Aware Scheme for Scheduling Big Data Applications with SAVANNA Hadoop
K. Reddy
,
Himansu Das
,
D. S. Roy
2017
Corpus ID: 57310724
2017
2017
Benchmarking Harp-DAAL: High Performance Hadoop on KNL Clusters
Langshi Chen
,
Bo Peng
,
+10 authors
Judy Qiu
IEEE International Conference on Cloud Computing
2017
Corpus ID: 540588
Data analytics is undergoing a revolution in many scientific domains, and demands cost-effective parallel data analysis…
Expand
2014
2014
Use of Big Data technology in Vehicular Ad-hoc Networks
Punam Bedi
,
Dr Vinita Jindal
International Conference on Advances in Computing…
2014
Corpus ID: 10561515
Big Data technology is becoming ubiquitous and depicting key attention of researchers in almost all areas. VANET is a special…
Expand
2014
2014
Research of the FP-Growth Algorithm Based on Cloud Environments
Lijuan Zhou
,
Xiang Wang
Journal of Software
2014
Corpus ID: 18129180
The emergence of cloud computing solves the problems that traditional data mining algorithms encounter when dealing with large…
Expand
2013
2013
Verification and Validation of MapReduce Progra m Model for Parallel Support Vector Machine Alg orithm on Hadoop Cluster
M. Kiran
,
Amresh Kumar
,
Saikat Mukherjee
,
G. Prakash
2013
Corpus ID: 212608502
We currently live in the data age. It’s not easy to measure the total volume of structured and unstructured data that require…
Expand
2012
2012
Complexity Measures for Map-Reduce, and Comparison to Parallel Computing
Ashish Goel
,
Kamesh Munagala
arXiv.org
2012
Corpus ID: 15713446
The programming paradigm Map-Reduce and its main open-source implementation, Hadoop, have had an enormous impact on large scale…
Expand
2011
2011
Performance Analysis of Hadoop for Query Processing
T. Wlodarczyk
,
Yi Han
,
Chunming Rong
IEEE Workshops of International Conference on…
2011
Corpus ID: 16469083
Query processing using mostly various NoSQL languages becomes a significant application area for Hadoop. Despite significant work…
Expand
2009
2009
Data Security Model for Cloud Computing
Yue Dai
2009
Corpus ID: 59792813
With the development of cloud computing, Data security becomes more and more important in cloud computing. This paper analyses…
Expand