Muhammad faisal
@muhammadfaisal20
I benchmark and refine LLMs while building reliable backend evaluation systems.
What I'm looking for
At Next AI, I train and refine next-generation LLMs using human conversations, improving GPT-generated web application output quality by 45% through iterative alignment feedback.
I evaluate AI-generated applications end to end, from multi-turn webpage generation to SWE-Bench code fixes in Python and Java. I’ve manually refactored generated frontend and backend code to reach 98% prompt compliance and functionality, and identified 50+ critical failure modes across Anthropic, Gemini, and OpenAI models.
I also build automated unit test suites, backend microservices, and Dockerized evaluation environments for reproducible AI code-fix pipelines.
Experience
Work history, roles, and key accomplishments
AI Engineer
Next AI
Aug 2024 - Present (2 years)
Trained next-generation LLMs (GPT-5) using human conversations, improving output quality by 45% through iterative alignment feedback. Conducted multi-turn evaluations and benchmarked LLM performance on SWE-bench, enhancing autonomous coding accuracy by 40%.
Education
Degrees, certifications, and relevant coursework
University of Management and Technology
Bachelor of Science, Computer Science
2021 - 2025
Pursuing a Bachelor of Science in Computer Science, with coursework in software engineering, web development, and cloud computing.
Availability
Location
Authorized to work in
Job categories
Skills
Interested in hiring Muhammad?
You can contact Muhammad and 90k+ other talented remote workers on Himalayas.
Message MuhammadGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
