Yashwanth Goud Matta
@yashwanthgoudmatta
I build high-scale AI infrastructure for LLM inference, RLHF, and distributed GPU systems.
What I'm looking for
At Scale AI, I architect enterprise RLHF and human-feedback systems processing 15B+ monthly inference requests, reducing inference costs by 29% while improving LLM alignment quality by 34%. I also build governance, multimodal processing, personalization, and workflow platforms using Java, Spring Boot, Kafka, Kubernetes, and distributed GPU infrastructure.
Previously at NVIDIA and Meta, I built GPU-orchestrated training and inference platforms, real-time control planes, and high-throughput ads ranking services. My work spans low-latency microservices, event-driven systems, Kubernetes, RAG, model serving, and production AI observability.
Experience
Work history, roles, and key accomplishments
Architected enterprise RLHF platform processing 15B+ monthly inference requests, reducing inference costs by 29% and improving LLM alignment quality by 34%. Engineered low-latency AI orchestration services achieving p95 latency under 350ms and built human-in-the-loop annotation systems supporting 420K+ reviewers.
Education
Degrees, certifications, and relevant coursework
University of Colorado Denver
Master of Science, Business Analytics
Pursued a Master's in Business Analytics at the University of Colorado Denver.
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Job categories
Skills
Interested in hiring Yashwanth Goud?
You can contact Yashwanth Goud and 90k+ other talented remote workers on Himalayas.
Message Yashwanth GoudGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
