This project implements a modern data engineering pipeline using Databricks, PySpark, DBT, and Delta Live Tables. It follows the Medallion Architecture, supports realtime data ingestion with Autoloader, and models data with fact and dimension tables, including Slowly Changing Dimensions (SCD Type 2), all orchestrated in a scalable cloud environment
pyspark databricks data-build-tool delta-lake streaming-pipeline delta-live-tables databricks-unity-catalog dimensional-data-modeling databricks-autoloder
-
Updated
Jul 15, 2025