
Raj Srujan Jalem
@rajsrujanjalem
I build and scale low-latency ML platforms across edge and cloud environments.
What I'm looking for
At Vimaan, I architect the core ML infrastructure and lead a team of nine engineers building high-performance Edge AI systems. My pipeline optimization reduced inference latency by 70% while processing 1–2 TB of multi-camera video data each day.
I've built distributed inference systems with ROS, Ray, Triton Inference Server, TensorRT, and dynamic batching, addressing edge bandwidth, packet loss, GPU cost, and ultra-low-latency requirements. I also created an LLMOps observability stack with OpenSearch for distributed logging, microservice health monitoring, and automated root-cause analysis.
Previously, I reduced annotation costs by about 40% at Alectio, cut API latency by 70% and accelerated deployments 3x at Fasal, and improved computer-vision throughput by 50% at Utopia Global. I enjoy taking AI platforms from prototype to reliable production systems while mentoring engineers and leading cross-functional teams.
Experience
Work history, roles, and key accomplishments
Manager, MLOps
Vimaan
Aug 2022 - Present (4 years 1 month)
Architected core ML infrastructure and managed a team of 9 engineers, achieving a 70% reduction in inference latency. Built a distributed, on-premise Edge AI pipeline using ROS streaming to process 1-2 TB of daily multi-camera video data.
ML Backend Engineer
Alectio
Jan 2022 - Aug 2022 (7 months)
Developed the Alectio Model Library to standardize model integration, automated training, and active-learning workflows. Built automated labeling pipelines that reduced customer data annotation costs by ~40%.
AI Data Engineer
Fasal
Mar 2021 - Jan 2022 (10 months)
Set up automated forecasting and continuous training workflows on Kubernetes using Kubeflow and MLflow. Migrated monolithic pipelines to a microservice architecture and implemented a Feature Store, reducing API latency by 70%.
Data Science Engineer
Utopia Global
Jan 2019 - Mar 2021 (2 years 2 months)
Trained custom U-Net models for super-resolution image enhancement, doubling the output resolution quality. Refactored production vision pipelines to increase processing throughput by 50%.
Staff MLOps Engineer
Vimaan
Resolved edge bandwidth bottlenecks via a dynamic frame-sampling module, using upstream object detection to trigger selective 4K image acquisition. Tuned ROS 1 and ROS 2 network configurations using Wireshark to diagnose packet loss, ensuring stable, high-throughput video streaming across edge nodes.
Senior MLOps Engineer
Vimaan
Scaled server-side inference using Ray and Triton Inference Server, applying TensorRT quantization and mixed-precision optimization alongside dynamic batching to cut GPU compute cost while sustaining ultra-low latency. Built an LLMOps observability stack using OpenSearch to aggregate distributed logs, monitor microservice health, and automate root-cause analysis.
Education
Degrees, certifications, and relevant coursework
IIIT Naya Raipur
Bachelor of Technology, Computer Science
2015 - 2019
Bachelor of Technology in Computer Science from IIIT Naya Raipur, completed in 2019.
Availability
Location
Authorized to work in
Portfolio
github.com/rajsrujanSalary expectations
Skills
Interested in hiring Raj Srujan?
You can contact Raj Srujan and 90k+ other talented remote workers on Himalayas.
Message Raj SrujanGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
