At Accenture Solutions, I build and maintain cloud data systems for the Research and Compliance project, managing large datasets from diverse sources for daily operations and client reporting.
I ingest data through SIP jobs scheduled with Autosys and land raw files in Azure Data Lake Storage. I import domain and mainframe data into Azure Data Lake, then convert it into governed Delta tables in Azure Databricks.
I develop and optimize PySpark, Spark SQL, Python, and DataFrame workloads, delivering up to 60% faster aggregations. Through partitioning and file compaction, I reduced query execution time by 50%.
I enforce schemas, data validation, and quality checks before loading curated outputs into Azure Synapse, supporting 98% reporting accuracy and 40% faster data retrieval. I also manage production notebook jobs and version-controlled code in GitHub.

