Netflix Data Engineering ETL Pipeline
This project is a real-world data engineering ETL pipeline built using a Kaggle Netflix dataset. The goal is to demonstrate how raw, messy data can be validated, cleaned, verified, and prepared for downstream analytics or data warehouse loading. The pipeline follows professional data engineering practices such as raw vs cleaned data separation, schema and quality validation, modular ETL design, and early failure on bad data. 🎯 What This Project Does Outputs (Saved Plots) Run
Netflix Data Engineering ETL Pipeline Read More »



