Type something to search...

Data Engineering

Blog Posts (1)

Article

Building a Scalable ETL Pipeline with AWS Glue (CSV to Parquet + Partitioning)

A hands-on walkthrough of a serverless ETL pipeline with AWS Glue, PySpark and Athena: raw CSV into partitioned Parquet for fast queries at scale.

08 Apr, 2026 Read