At Genpact, I built lakehouse architecture using AWS Iceberg and Apache Spark for versioned data analytics and time-travel queries. I developed PySpark ETL pipelines on AWS Glue and implemented data governance with AWS Glue Catalog, Lake Formation, and IAM policies.
Across financial, automotive, textile, and retail projects, I developed data pipelines using AWS and Azure services, Databricks, and PySpark. My work included ingesting and transforming data, optimizing analytics workflows, and creating Tableau and Power BI dashboards.

