
眠眠 姜
@0005960
I evaluate Chinese AI outputs for safety, accuracy, instruction following, and content quality.
What I'm looking for
I review Chinese-language AI outputs for safety, accuracy, relevance, clarity, completeness, and instruction following. My work combines structured evaluation criteria with clear, concise feedback.
In educational content review, I checked materials, exercises, answer keys, and explanations before delivery to students. I identified incorrect answers, logical inconsistencies, unclear explanations, and formatting issues while applying predefined standards across high-volume content.
Through independent AI safety evaluation projects, I assessed prompts and responses for intent, adversarial framing, dual-use potential, escalation, severity, and risk category. I classified privacy, fraud, cybersecurity, harassment, and sensitive-information risks and recommended Allow, Partial, or Refuse outcomes with documented rationale.
I compare outputs from ChatGPT, Claude, Gemini, and DeepSeek, identify hallucinations and unsupported claims, and refine ambiguous prompts into practical review checkpoints. I also design quality-control and formatting workflows for Word-based educational documents, using AI to improve efficiency while manually verifying final results.
Experience
Work history, roles, and key accomplishments
LLM Response Quality Evaluation
Independent Project
Jan 2026 - Present (8 months)
Compared outputs from ChatGPT, Claude, Gemini, and DeepSeek for accuracy, relevance, completeness, clarity, instruction following, and safety. Identified hallucinations, unsupported claims, and missed instructions; refined prompts and converted ambiguous requests into review checkpoints.
Document Quality and Workflow Automation
Independent Project
Jan 2026 - Present (8 months)
Designed review and formatting workflows for Word-based educational documents, defined quality-control rules, and identified recurring AI output errors requiring human verification.
Chinese AI Safety Evaluation Practice
Independent Project
Jan 2026 - Present (8 months)
Evaluated Chinese-language prompts and model responses using a structured framework covering intent, risk category, adversarial framing, dual-use potential, escalation, severity, decision, and rationale. Classified safety risks and produced concise rationales with recommended outcomes.
Content Review and Teaching Assistant
To Fill Organization
Reviewed educational materials, exercises, answer keys, and explanations for accuracy, clarity, consistency, and formatting quality. Applied predefined standards across high-volume content and documented correction needs.
Education
Degrees, certifications, and relevant coursework
眠眠 hasn't added their education
Don't worry, there are 90k+ talented remote workers on Himalayas
Tech stack
Software and tools used professionally
Availability
Location
Authorized to work in
Salary expectations
Job categories
Interested in hiring 眠眠?
You can contact 眠眠 and 90k+ other talented remote workers on Himalayas.
Message 眠眠Get matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
