shashank shukla
@shashankshukla
Data Engineer specializing in scalable, validated ETL/ELT pipelines using Python, PySpark, Azure, Databricks and GCP.
What I'm looking for
I’m a Data Engineer with 4 years of experience designing scalable ETL/ELT pipelines and ensuring high-quality, validated data at scale. I focus on optimizing large-scale data systems and maintaining data integrity across 2TB+ daily pipelines.
In freelance work, I executed end-to-end migration of 10TB+ data from SQL Server to Azure Synapse using Azure Data Factory and Azure Databricks across multiple operational domains. I built CDC-based incremental pipelines that land raw data as Parquet in the bronze layer, reducing data processing volume by ~60–70%.
I implement Medallion Architecture (Bronze/Silver/Gold) with scalable PySpark transformations, including deduplication, schema enforcement, null handling, and cross-domain joins—cutting downstream query time by ~25–40%. I also apply business rule validations and data quality frameworks across pipeline stages to ensure accuracy, completeness, and consistency.
Earlier, as a Junior Analyst (Data Engineer), I built production ETL pipelines on Azure Data Factory and Databricks for 2TB+ daily data and supported hybrid banking migrations to Azure Synapse with governance compliance. I designed incremental partition-based processing to reduce daily processing time by 30–35%, and created event-driven ELT pipelines on GCP using GCS notifications to reduce latency by ~60–70%.
Experience
Work history, roles, and key accomplishments
Data Engineer (Freelance)
Anmol Gau Dugdhshaala
Feb 2025 - Present (1 year 5 months)
Executed end-to-end migration of 10TB+ data from SQL Server to Azure Synapse using Azure Data Factory and Azure Databricks across multiple operational domains. Built CDC-based incremental PySpark pipelines with Medallion architecture and implemented data quality validations, improving downstream query time and processing efficiency.
Junior Analyst (Data Engineer)
Course5 Intelligence Ltd.
Jul 2022 - Jan 2025 (2 years 6 months)
Built production ETL pipelines using Azure Data Factory and Databricks, processing 2TB+ of daily data from SQL Server and file systems. Led large-scale banking data migrations to Azure Synapse and developed incremental and event-driven ELT pipelines on Azure and GCP to reduce processing time and data latency.
Education
Degrees, certifications, and relevant coursework
Vasantdada Patil Pratishthan's College of Engineering and Visual Arts
Bachelor of Engineering, Computer Engineering
2018 - 2022
Completed a Bachelor of Engineering in Computer Engineering from Vasantdada Patil Pratishthan's College of Engineering and Visual Arts (2018–2022).
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Job categories
Skills
Interested in hiring shashank?
You can contact shashank and 90k+ other talented remote workers on Himalayas.
Message shashankGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
