I'm currently delivering data ingestion and transformation pipelines for Citi Bank at LTIMindtree, including UDM transaction ingestion that enabled analysis of Citi customer transactions on the SpaceX IPO launch day.
I've optimized Spark pipelines, implemented CDC logic, and built onboarding workflows from raw files through staging and standardization layers with data-quality checks. I also lead small teams, coordinate with QA, and automate reporting and manual development activities using PySpark, Python, and shell scripts.
Previously at ZS Associates, I owned a Spravato Digital Asset pipeline for Janssen from requirements gathering through production, leading a team of four and working with internal data vendors and external partners. I also developed AWS-based ETL workflows, SQL transformations, Redshift tables, and Tableau dashboard data sources.
At Wipro, I enhanced PySpark ETL pipelines for Charles Schwab, supporting new data sources and historical Hive-based tracking. Across BFSI and healthcare projects, I focus on reliable, scalable data delivery and practical automation.
