At IQVIA, I designed, built, and optimized ETL/ELT pipelines with Python and PySpark, transforming structured and unstructured data for downstream analytics.
I automated file ingestion for XLSX, CSV, and ZIP formats, reducing manual processing time by ~40% through automation and schema validation.
I implemented data quality checks and Pytest unit tests, achieving >90% code coverage on core transformation modules. I also worked with GitHub Actions on automated testing and deployment pipelines.
On my Medallion Architecture personal project, I developed an Azure pipeline spanning ADLS Gen2, Azure Data Factory, Databricks, and Synapse. I connected the Gold layer to Power BI for interactive reporting.

