Freelance AI Evaluation Engineer (Python/Full-Stack) opportunity involves creating coding test cases, reviewing and refining realistic coding tasks, writing functional tests, analyzing AI failures, and iterating based on feedback from expert QA reviewers.
Requirements
- Degree in Computer Science, Software Engineering or related fields
- 5+ years in software development, primarily Python (pytest, async/await, subprocess, file operations)
- Background in Full-Stack development, with an equal focus on building React-based interfaces and robust Back-end systems
- Experience writing tests (functional, integration – not just running them)
- Docker containers (running evaluations locally in containers)
- CI/CD understanding (GitHub Actions as a user: triggers, labels, reading results)
- English proficiency - B2
Benefits
- Paid hourly rate of up to $50 per hour equivalent
- Flexible work schedule
- Variety of projects with different earning levels based on scope, complexity, and required expertise
