At Oriserve, I fine-tuned speech and ASR models for Indian multilingual telephony, including Whisper, Qwen-ASR, and NVIDIA NeMo. I applied a calm-whisper technique that reduced Whisper ASR hallucinations by 90% and used weighted cross entropy to reduce WER by 15%.
I also built production GenAI solutions for EF Education First and TVS Motor using LLMs, RAG, FastAPI, and Pinecone. In research at the Machine Learning and Geocomputing Lab, IIT ISM Dhanbad, I trained diffusion models on OpenFWI seismic data and achieved 86% accuracy in seismic FWI.

