Project summary · Microsoft Fabric · PySpark · T-SQL · Power BI
A full medallion lakehouse in Fabric — raw landing, PySpark cleanup, a T-SQL star schema in the Warehouse, and a deployment pipeline that promotes it all between workspaces.
Job descriptions kept asking for Fabric plus a real promotion path between environments — not a single workspace with everything dumped in it. This project builds the whole stack the way a team would run it, including source control and CI/CD.


The dataset and the overall pattern come from Ansh Lamba's Fabric tutorial, which is attributed in the repo README. The pipeline build, the cleanup logic, the data model and the CI/CD setup are my own implementation.