At Salesforce, I reduced Spark job failures by 60% and improved distributed processing throughput by 35% across batch workloads. I also deployed 20+ production data ingestion workflows, improving data quality through schema validation and incremental loading.
At Informatica, I built ETL/ELT pipelines for Snowflake, Databricks, and Microsoft Fabric, reducing runtimes by 70–80%. I automated orchestration with event-driven workflows and REST APIs, eliminating 90% of manual intervention.
On my Customer 360 Data Platform project, I built a fault-tolerant pipeline with AWS S3 and Delta Lake that reduced incremental batch size by 99%. I also created a Databricks SQL analytics mart for customer segmentation and churn-risk analysis.

