Skip to main content
📘 Complete PDF

PySpark Certification Practice Test 07

Find The Problem In The Code

Question 1 of 20
📋 Multiple Choice
Question 121
A developer writes: ```python df = spark.read.parquet("/sales") result = df.collect() print(len(result)) ``` The dataset contains 2 billion records. What is the biggest problem?
A. Driver Out Of Memory Risk
B. Missing Cache
C. Missing Broadcast
D. Missing Repartition
Answer
✅ A. Driver Out Of Memory Risk
Explanation
### Explanation collect() moves all data to the Driver.
Answered: 0 / 20
Question Navigator (Test 07)0% Answered