Recent questions tagged spark

0 votes
1 answer

Spark - load CSV file as DataFrame?

Sep 25, 2018 in Big Data Hadoop by digger
• 26,740 points
7,476 views
0 votes
1 answer
0 votes
1 answer

What happens to RDD when one of the nodes goes down?

Sep 3, 2018 in Apache Spark by Shubham
• 13,490 points
2,546 views
0 votes
1 answer

Does Spark provide the storage layer too?

Sep 3, 2018 in Apache Spark by Shubham
• 13,490 points
2,157 views
0 votes
1 answer

Functions of Spark SQL?

Sep 3, 2018 in Apache Spark by Meci Matt
• 9,460 points
2,271 views
0 votes
1 answer

Languages supported by Apache Spark?

Sep 3, 2018 in Apache Spark by Meci Matt
• 9,460 points
8,202 views
0 votes
1 answer

How to connect Amazon RedShift in Apache Spark?

Aug 22, 2018 in AWS by datageek
• 2,540 points
8,713 views
+2 votes
3 answers
0 votes
2 answers

Which cluster type should I choose for Spark?

Aug 21, 2018 in Apache Spark by Shubham
• 13,490 points
2,928 views
0 votes
1 answer
0 votes
2 answers

Which of these will vanish: Flink vs Spark?

Aug 10, 2018 in Big Data Hadoop by Omkar
• 69,180 points
2,098 views
0 votes
2 answers
0 votes
1 answer

What makes Spark faster than MapReduce?

Jul 27, 2018 in Apache Spark by Neha
• 6,300 points
2,423 views
0 votes
1 answer

PySpark Config ?

Jul 26, 2018 in Apache Spark by shams
• 3,670 points
1,529 views
+1 vote
1 answer
0 votes
1 answer

A strange spark ERROR on AWS EMR

Jul 13, 2018 in AWS by Luke cage
• 360 points
2,377 views
+1 vote
8 answers

How to print the contents of RDD in Apache Spark?

Jul 6, 2018 in Apache Spark by Shubham
• 13,490 points
65,628 views
0 votes
2 answers

How to use RDD filter with other function?

Jul 5, 2018 in Apache Spark by Shubham
• 13,490 points
10,790 views
0 votes
1 answer

How to add third party java jars for use in PySpark?

Jul 4, 2018 in Apache Spark by Shubham
• 13,490 points
9,662 views
0 votes
1 answer
0 votes
1 answer
+1 vote
1 answer

map vs mapValues in Spark

Jun 29, 2018 in Apache Spark by Shubham
• 13,490 points
17,445 views
+1 vote
3 answers

Which cluster type should I choose for Spark?

Jun 27, 2018 in Apache Spark by Shubham
• 13,490 points
2,677 views
0 votes
1 answer

Which is better in term of speed, Shark or Spark?

Jun 26, 2018 in Apache Spark by Shubham
• 13,490 points
1,603 views
0 votes
1 answer

Spark Driver roles

Jun 21, 2018 in Apache Spark by shams
• 3,670 points
1,652 views
0 votes
2 answers

Parquet Files Advantages

Jun 21, 2018 in Apache Spark by Data_Nerd
• 2,390 points
2,980 views
0 votes
2 answers

map() and flatmap()

Jun 20, 2018 in Apache Spark by Ashish
• 2,650 points
1,944 views
0 votes
1 answer

Spark standalone client mode

Jun 20, 2018 in Apache Spark by shams
• 3,670 points
1,888 views
0 votes
1 answer

Ways to create RDD in Apache Spark

Jun 19, 2018 in Apache Spark by Shubham
• 13,490 points
4,899 views
0 votes
3 answers

Lineage Graph in Spark

Jun 19, 2018 in Apache Spark by Data_Nerd
• 2,390 points
13,493 views
0 votes
1 answer
0 votes
1 answer

How RDD persist the data in Spark?

Jun 18, 2018 in Apache Spark by kurt_cobain
• 9,350 points
2,127 views
0 votes
1 answer

What do we mean by an RDD in Spark?

Jun 18, 2018 in Apache Spark by kurt_cobain
• 9,350 points
4,846 views
0 votes
1 answer
0 votes
1 answer
0 votes
1 answer

Persistence Levels in Spark

Jun 8, 2018 in Apache Spark by Data_Nerd
• 2,390 points
6,939 views
0 votes
1 answer

What is Shark?

Jun 8, 2018 in Apache Spark by shams
• 3,670 points
1,731 views
+1 vote
1 answer

Kafka Feature

Jun 7, 2018 in Apache Spark by shams
• 3,670 points
2,731 views
0 votes
1 answer

SQLInterpreter in Spark

Jun 7, 2018 in Apache Spark by shams
• 3,670 points
1,429 views
0 votes
1 answer
0 votes
1 answer

Parquet File

Jun 4, 2018 in Apache Spark by shams
• 3,670 points
1,708 views