At Celebal Technologies Pvt. Ltd., I designed a real-time ingestion pipeline with PySpark Structured Streaming on Databricks, bringing API and Kafka data into Delta Lake. I also developed incremental data solutions for more than 15 TB to support analytics and Data Science model training and fine-tuning.
At CloudGritz Technologies Pvt. Ltd., I developed ETL pipelines to migrate more than 200 SQL Server tables to Amazon Redshift, processing over 2 TB of daily incremental data. I also worked on a Databricks real-time data lake, optimizing query performance and strengthening pipeline reliability.

