Skip to main content
Ankur RoyAR
Looking for a job

Ankur Roy

@ankurroy

I evaluate and train LLMs through RLHF, multi-turn reasoning critique, and rigorous factuality verification.

India
Message

What I'm looking for

I'm looking to contribute rigorous LLM evaluation, RLHF, factuality verification, and multi-turn reasoning analysis in a remote AI training or data quality role.

I've evaluated and trained LLM responses for Outlier.ai, ranking multi-turn outputs for truthfulness, instruction adherence, conciseness, and tone. I write detailed rationales that identify logical fallacies, rubric violations, mathematical inaccuracies, and hallucinations while sustaining a 95%+ quality audit rating.

Across Atlas Data Annotation and Remotasks, I've labeled dialogue data, audited conversational context, evaluated chain-of-thought outputs, and conducted side-by-side preference evaluations. I bring engineering rigor, structured critical thinking, and high-velocity fact-checking to improving generative AI training data.

Experience

Work history, roles, and key accomplishments

Outlier.ai logoOU
Current

AI Trainer & Evaluation Specialist

Jan 2024 - Present (2 years 8 months)

Evaluate, rank, and rate multi-turn LLM candidate responses across truthfulness, instruction adherence, conciseness, and tone. Author justification rationales and conduct source-level verification to detect hallucinations.

Education

Degrees, certifications, and relevant coursework

HT

Heritage Institute of Technology

B.Tech, Civil Engineering

Graduated with a B.Tech in Civil Engineering in 2025.

Tech stack

Software and tools used professionally

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan