At Cognizant Technology Solutions, I build and maintain batch ETL/ELT pipelines that integrate databases, APIs, files, and cloud storage into reliable analytical data solutions.
I use Azure Data Factory, Azure Databricks, Python, SQL, and PySpark to transform large datasets through cleansing, deduplication, validation, enrichment, aggregation, and incremental processing.
I've designed fact and dimension tables, star schemas, and SCD Type 1 and Type 2 models for reporting and downstream data consumption. I also implement Bronze, Silver, and Gold data layers with ADLS Gen2 and Delta Lake.
I monitor production pipelines, resolve data-quality and performance issues, and optimize Spark workloads with partitioning, caching, broadcast joins, and query optimization. I work with analysts, QA engineers, and development teams in Agile environments to deliver maintainable data solutions.

