Ардней Воланду
@0004956
I evaluate large language models, uncover failure modes, and improve results through systematic prompting and testing.
What I'm looking for
I've spent more than two years independently researching and working intensively with ChatGPT, DeepSeek, Google Gemini, Claude, Grok, and Microsoft Copilot.
I compare LLMs on complex tasks, assess response quality, reasoning, consistency, context retention, and failure modes, and identify contradictions, hallucinations, and context loss. I design complex prompts and multi-step interaction scenarios, then iteratively refine task formulations to produce more reliable and useful results.
From a professional home workstation, I handle long conversational contexts and large volumes of information for sustained remote AI evaluation, testing, training, and research work.
Experience
Work history, roles, and key accomplishments
Advanced LLM Specialist / AI Evaluator
Independent
Jan 2024 - Jan 2026 (2 years)
Independent LLM specialist with over 2 years of intensive experience researching and evaluating modern AI systems, focusing on comparative analysis, complex task solving, and systematic investigation of model capabilities and limitations.
Education
Degrees, certifications, and relevant coursework
Ардней hasn't added their education
Don't worry, there are 90k+ talented remote workers on Himalayas
Availability
Location
Authorized to work in
Job categories
Interested in hiring Ардней?
You can contact Ардней and 90k+ other talented remote workers on Himalayas.
Message АрднейGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
