At Cyient Ltd, I build scalable AI-ready data pipelines using Python, SQL, PySpark, Databricks, AWS Glue, and DBT for ML, analytics, and GenAI use cases.
I've migrated legacy on-premise platforms to Snowflake and BigQuery, created ML-ready feature tables and curated data layers, and optimized large-scale workloads through Spark tuning, SQL tuning, partitioning, clustering, and indexing.
I automate validation, profiling, cleansing, transformation, and batch processing so teams can rely on clean, feature-ready datasets for AI, NLP, reporting, and machine learning applications.
Previously at Metro Labs Pvt Ltd, I developed Python ETL pipelines, data models, curated datasets, and production monitoring for analytics and AI workloads. I also implement secure cloud-native architecture, CI/CD, monitoring, alerting, failure handling, and SLA tracking across GCP and AWS.
