Skip to main content
Ankur PatelAP
Looking for a job

Ankur Patel

@ankurpatel

Senior Data Engineer (5.8 yrs) — Azure Databricks, PySpark, Kafka, Delta Lake. Lakehouse, real-time streaming, CI/CD, and production GenAI/RAG

India
Message

What I'm looking for

I am looking for opportunities that allow me to leverage my data engineering skills in a collaborative environment, focusing on innovative projects that drive business success.

Senior Data Engineer, 5.8 years, building production data platforms on Azure Databricks — lakehouse, real-time streaming, CI/CD, and GenAI that ships to real users.

Currently at DCM Shriram Limited, a listed Indian chemicals manufacturer, where I own the end-to-end Data & AI platform and Azure DevOps CI/CD. Two plant-optimization platforms I built end-to-end have delivered $1.2M (₹10.8 Cr) in measured savings, and re-architecting the data platform cut Azure cloud spend 65% in under 6 months — root-caused a mis-set blob access tier driving 99.5% of storage operation cost, capped Kafka to 1 eCKU, retuned pipeline triggers, and right-sized clusters and App Service plans. I set technical direction for a cross-functional team of 4 (2 Data Engineers, 2 Data Scientists).

Before that, 5 years as a Data & Analytics Engineer on multi-year engagements with international clients across e-commerce, retail, and healthcare — Kafka + Databricks streaming, Snowflake + dbt warehouses, CDP implementations, and HIPAA-compliant healthcare pipelines.

What I build:

- Config-driven streaming frameworks on Databricks — one Lakeflow Spark Declarative Pipeline serving every real-time use case through pure configuration; 1/3 the infrastructure cost of the setup it replaced, zero repeated failures, 1-second event latency

- Production GenAI / RAG on Azure AI Search and Azure OpenAI — hybrid BM25 + vector retrieval, semantic reranking, metadata filtering, and source citations across 70 GB of documentation

- Lakehouse platforms on Delta Lake with Unity Catalog and a Databricks Asset Bundle monorepo across DEV / TEST / PROD

- Real-time Kafka pipelines and medallion architectures processing millions of daily events

- Full-stack data applications — FastAPI + PostgreSQL fronting Pyomo optimization models (MILP / nonlinear)

Stack

Daily: Python · SQL · PySpark · Azure Databricks · Delta Lake · Delta Live Tables · Unity Catalog · Databricks Asset Bundles · Kafka · FastAPI · Docker · Azure DevOps

Regular: dbt · Snowflake · Azure Data Factory · Synapse · Terraform · PostgreSQL · Azure Functions · React

GenAI: Azure AI Search · Azure OpenAI · LangChain · FAISS · Pyomo (MILP / nonlinear)

Experience

Work history, roles, and key accomplishments

DL
Current

Senior Data Engineer

DCM Shriram Limited

Oct 2025 - Jul 2026 (9 months)

Own the end-to-end Data & AI platform on Azure Databricks — real-time streaming, lakehouse, CI/CD, and production GenAI/RAG serving multiple chemical manufacturing plants. Re-architecture cut Azure cloud spend 65% in under 6 months; two end-to-end optimization platforms delivered $1.2M in measured savings. Set technical direction for a team of 4 (2 DE, 2 DS) and own Azure DevOps CI/CD.

SE

Data Engineering Consultant

Self-employed

Aug 2020 - Aug 2025 (5 years)

Multi-year engagements with international clients across e-commerce, retail, and healthcare. Architected 30+ production pipelines processing 10M+ records daily on Databricks, PySpark, and Azure. Built Kafka + Databricks Structured Streaming platforms (client-reported 28% revenue lift), Snowflake + dbt customer data platforms, and HIPAA-compliant healthcare pipelines (500K+ daily records).

Education

Degrees, certifications, and relevant coursework

VC

Vishwakarma Government Engineering College

B.Tech. in Information & Technology Engineering, Information & Technology Engineering

2017 - 2020

Grade: 8.60 CGPA

Activities and societies: Activities and societies: - Engaged in coding hackathons and online challenges - Joined seminars on emerging IT technologies - Attended workshops on data analytics and machine learning - Participated in inter-college coding contests - Participated in open-source project contributions during college

Graduated with a B.Tech in Information Technology.

Focused on programming, databases, data engineering, and data analytics.
Built foundations through coursework in algorithms, distributed systems, and database design — later shaped into a full career in production data engineering.

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan