At LTM, I built AWS Glue ETL pipelines for a healthcare client's Provider Tiering Data Platform, ingesting 54M+ provider records from Amazon RDS into S3 datasets. I implemented partner-level parallel processing that made batch processing 14x faster.
I also implemented CDC-based incremental loads and developed metadata-driven SQL templates that let new partners onboard through configuration alone. I added data quality checks, execution logging, SNS alerts, and S3-hosted reports to support reliable downstream use.
On my Data Lakehouse Analytics Platform project, I built metadata-driven PySpark pipelines and a Bronze/Silver/Gold architecture on Delta Lake, with schema drift tracking and SCD Type 2. I’m a Databricks Certified Data Engineer Associate and an AWS Certified Cloud Practitioner.

