At Abbott, I designed and optimized Spark and PySpark jobs in Databricks, improving data processing efficiency by 30%. I also developed dimensional data models and implemented data quality and lineage frameworks across production pipelines.
I’ve built and maintained data pipelines at GoldmanSachs, IBM, and General Electric, using Airflow to orchestrate workflows and Kafka or Kinesis for streaming. At IBM, I managed more than 50 Airflow DAGs supporting data ingestion and transformation.

