I’m looking for a Data Engineering role where I can build scalable data pipelines and work with Azure, Databricks, PySpark, SQL, and modern data platforms. I’m interested in solving large-scale data challenges, improving pipeline performance and reliability, and expanding my expertise in Spark, cloud technologies, streaming, and lakehouse architectures.

Sameer Kumar
@sameerkumar10
Data Engineer with 3 years of experience in Azure, Databricks, PySpark, SQL, ADF, ETL/ELT pipelines, Power BI, and data optimization.
What I'm looking for
I am a Data Engineer with around 3 years of experience building scalable data solutions using Azure Data Factory, Azure Databricks, PySpark, Python, SQL, Delta Lake, ADLS, and Power BI. In my current role as a Software Engineer II at MAQ Software, I work across the complete data pipeline lifecycle—from understanding business requirements and developing ingestion pipelines from scratch to data transformation, validation, optimization, monitoring, troubleshooting, and production support.
I have experience building batch and Spark Structured Streaming pipelines, processing multi-terabyte datasets, implementing Medallion (Bronze/Silver/Gold) architectures, and developing incremental/CDC-based ingestion solutions. I have also worked on schema validation, anomaly detection, data modeling, REST API integrations, and CI/CD using Azure DevOps, Git, ARM templates, and automated testing.
Some of my key achievements include reducing ingestion-related support issues by 30%, reducing dataset size by 50% through data model optimization, and improving delivery time by 30% through CI/CD automation. I have also developed 15+ Power BI reports and worked on an Azure OpenAI-based solution using prompt engineering.
I am particularly interested in Data Engineering, Big Data, distributed processing, Databricks, Apache Spark, cloud data platforms, lakehouse architectures, real-time data processing, and building reliable, high-performance data pipelines. I am looking for opportunities where I can solve large-scale data engineering problems, deepen my expertise in modern cloud and data technologies, and contribute to scalable production data platforms.
Experience
Work history, roles, and key accomplishments
Software Engineer II
MAQ Software
Mar 2026 - Present (7 months)
Architected end-to-end batch ETL pipelines on a Medallion lakehouse using Azure Data Factory, Databricks, and Delta Lake. Implemented CI/CD for data pipelines, reducing delivery time by 30% and improving report accuracy by 25%.
Software Engineer I
MAQ Software
Sep 2024 - Feb 2026 (1 year 5 months)
Built and optimized batch ETL pipelines in Azure Data Factory, integrating Python and PySpark notebooks for schema validation and anomaly detection. Engineered high-performance data models and configured row-level security in Power BI.
Associate Software Engineer
MAQ Software
Sep 2023 - Aug 2024 (11 months)
Delivered and enhanced 15+ Power BI reports aligned with business requirements. Implemented an Azure OpenAI-based solution for predictive column generation and led an Azure DevOps project migration.
Education
Degrees, certifications, and relevant coursework
GL Bajaj Institute of Technology and Management
Bachelor of Technology, Information Technology
2020 - 2024
Pursued a Bachelor of Technology in Information Technology from 2020 to 2024.
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Salary expectations
Social media
Job categories
Interested in hiring Sameer?
You can contact Sameer and 90k+ other talented remote workers on Himalayas.
Message SameerGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
