Which Data Structure is used in case of Map Reduce

I am working on Hadoop from last 5 months but I am wondering that which data structure Hadoop uses to store large datasets. Where can I find a detailed view of underlying data structures in Hadoop? or Can anyone explain me here?

Thanks in advance!

May 4, 2018 in Big Data Hadoop by Shubham
• 13,490 points • 1,786 views

1 answer to this question.

In case of Hadoop, HDFS is used as a storage platform and The underlying HDFS uses block as a storing units. HDFS does not care about what structure the files have. A MapReduce program simply gets the file data from HDFS as an input.

You can refer the below book to get a complete detailed information:

https://www.isical.ac.in/~acmsc/WBDA2015/slides/hg/Oreilly.Hadoop.The.Definitive.Guide.3rd.Edition.Jan.2012.pdf

Hope this will help you!

answered May 4, 2018 by nitinrawat895
• 11,380 points

Related Questions In Big Data Hadoop

0 votes

1 answer

Which data type is used to store the data in HBase table column?

Hey, Byte Array, Put p = new Put(Bytes.toBytes("John Smith")); All ...READ MORE

answered May 29, 2019 in Big Data Hadoop by Gitika
• 65,730 points • 2,519 views

0 votes

1 answer

What is the purpose of shuffling and sorting phase in the reducer in Map Reduce?

Hi@akhtar, Shuffle phase in Hadoop transfers the map output from ...READ MORE

answered Dec 20, 2020 in Big Data Hadoop by MD
• 95,460 points • 8,252 views

0 votes

1 answer

Which is the most preferable language for Hadooop Map-Reduce programs?

MapReduce is a programming model to perform ...READ MORE

answered Aug 4, 2018 in Big Data Hadoop by Neha
• 6,300 points • 4,262 views

0 votes

1 answer

Which side join is taken by default by hive? Map-side or Reduce-side?

The syntax for Map-side join and Reduce-side ...READ MORE

answered Dec 13, 2018 in Big Data Hadoop by Omkar
• 69,180 points • 1,438 views

–1 vote

1 answer

Hadoop dfs -ls command?

In your case there is no difference ...READ MORE

answered Mar 16, 2018 in Big Data Hadoop by kurt_cobain
• 9,350 points • 5,090 views

+1 vote

1 answer

Hadoop Mapreduce word count Program

Firstly you need to understand the concept ...READ MORE

answered Mar 16, 2018 in Data Analytics by nitinrawat895
• 11,380 points • 11,615 views

+2 votes

11 answers

hadoop fs -put command?

Hi, You can create one directory in HDFS ...READ MORE

answered Mar 16, 2018 in Big Data Hadoop by nitinrawat895
• 11,380 points • 112,659 views

0 votes

1 answer

Is there a way to copy data from one one Hadoop distributed file system(HDFS) to another HDFS?

The distributed copy command, distcp, is a ...READ MORE

answered Mar 22, 2018 in Big Data Hadoop by Ashish
• 2,650 points • 10,194 views

0 votes

1 answer

How is kafka used in big-data?

I can brief you the answer here ...READ MORE

answered Mar 26, 2019 in Big Data Hadoop by nitinrawat895
• 11,380 points • 1,182 views

0 votes

1 answer

How Impala is fast compared to Hive in terms of query response?

Impala provides faster response as it uses MPP(massively ...READ MORE

answered Mar 21, 2018 in Big Data Hadoop by nitinrawat895
• 11,380 points • 2,692 views

Subscribe to our Newsletter, and get personalized recommendations.

REGISTER FOR FREE WEBINAR

Thank you for registering Join Edureka Meetup community for 100+ Free Webinars each month JOIN MEETUP GROUP