At Outlier, I evaluate AI-generated responses for accuracy, logic, relevance, and instruction adherence. I compare outputs against project-specific rubrics and write rationales explaining their strengths and weaknesses.
At Remotasks, I annotated and quality-checked training data, and at Handshake I reviewed English, Swahili, and Spanish content for linguistic quality and meaning. My computer science background supports my evaluation of code responses, algorithms, and debugging logic.

