Deepanshu Yadav
@deepanshuyadav1
AI Research Intern building production RAG and real-time AI services with FastAPI on GCP.
What I'm looking for
I built a production RAG pipeline at Indixpert Technology, deploying Vertex AI embeddings with pgvector on Cloud Run and serving real-time AI responses through 15+ async REST endpoints.
To keep costs and latency down, I added a semantic caching layer to reduce redundant Gemini API requests and replaced Vertex AI Vector Search with pgvector on Cloud SQL.
I also implemented an auto-escalation pipeline where sentiment scoring triggers P1 ticket creation and agent handoff, designed for zero human triage.
Outside that work, I’ve shipped AI apps like Langalytics for conversational CSV analysis and a YOLOv8-based traffic management system for detection, tracking, and counting.
Experience
Work history, roles, and key accomplishments
AI Research Intern
Indixpert Technology
May 2026 - Jul 2026 (2 months)
Designed and deployed a production RAG pipeline using Vertex AI embeddings, pgvector, and Gemini 2.5 Flash on Cloud Run with FastAPI. Implemented semantic caching and auto-escalation pipeline with real-time sentiment scoring.
Education
Degrees, certifications, and relevant coursework
DIT University
Bachelor of Technology, Information Technology
2023 -
Grade: 7.36/10
Pursuing Bachelor of Technology in Information Technology with a CGPA of 7.36/10.
Availability
Location
Authorized to work in
Portfolio
my-portfolio-9q9o.onrender.comJob categories
Interested in hiring Deepanshu?
You can contact Deepanshu and 90k+ other talented remote workers on Himalayas.
Message DeepanshuGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
