PYSPARK ACADEMY • MODULE 03Reading Data
Read structured and semi-structured file formats, configure headers and delimiters, use inferSchema, and load enterprise Parquet files.
Beginner📚 8 Lessons⭐ 0 / 1420 XP⏳ In Progress
0 / 8 Lessons Completed (0%) LESSON 17🔓 Reading CSV Files
⏱ 2 Mins • ⭐ 150 XP • Beginner
LESSON 18🔒 Configuring CSV Parsing Options
⏱ 2 Mins • ⭐ 160 XP • Beginner
LESSON 19🔒 Handling CSV Headers with option()
⏱ 2 Mins • ⭐ 165 XP • Beginner
LESSON 20🔒 Automatic Schema Inference with inferSchema
⏱ 2 Mins • ⭐ 170 XP • Beginner
LESSON 21🔒 Reading JSON Datasets
⏱ 2 Mins • ⭐ 180 XP • Beginner
LESSON 22🔒 Reading Optimized Parquet Files
⏱ 2 Mins • ⭐ 190 XP • Beginner
LESSON 23🔒 Generic Ingestion with read.format()
⏱ 2 Mins • ⭐ 195 XP • Beginner
LESSON 24🔒 Production-Style Ingestion with Schema Merging
⏱ 3 Mins • ⭐ 210 XP • Beginner