Skip to main content
眠眠 姜
Looking for a job

眠眠

@0005960

I evaluate Chinese AI outputs for safety, accuracy, instruction following, and content quality.

China
Message

What I'm looking for

I'm looking for a role evaluating Chinese-language AI content, model responses, safety risks, or content quality using clear guidelines, structured reasoning, and human verification.

I review Chinese-language AI outputs for safety, accuracy, relevance, clarity, completeness, and instruction following. My work combines structured evaluation criteria with clear, concise feedback.

In educational content review, I checked materials, exercises, answer keys, and explanations before delivery to students. I identified incorrect answers, logical inconsistencies, unclear explanations, and formatting issues while applying predefined standards across high-volume content.

Through independent AI safety evaluation projects, I assessed prompts and responses for intent, adversarial framing, dual-use potential, escalation, severity, and risk category. I classified privacy, fraud, cybersecurity, harassment, and sensitive-information risks and recommended Allow, Partial, or Refuse outcomes with documented rationale.

I compare outputs from ChatGPT, Claude, Gemini, and DeepSeek, identify hallucinations and unsupported claims, and refine ambiguous prompts into practical review checkpoints. I also design quality-control and formatting workflows for Word-based educational documents, using AI to improve efficiency while manually verifying final results.

Experience

Work history, roles, and key accomplishments

IP

LLM Response Quality Evaluation

Independent Project

Jan 2026 - Present (8 months)

Compared outputs from ChatGPT, Claude, Gemini, and DeepSeek for accuracy, relevance, completeness, clarity, instruction following, and safety. Identified hallucinations, unsupported claims, and missed instructions; refined prompts and converted ambiguous requests into review checkpoints.

IP

Document Quality and Workflow Automation

Independent Project

Jan 2026 - Present (8 months)

Designed review and formatting workflows for Word-based educational documents, defined quality-control rules, and identified recurring AI output errors requiring human verification.

IP

Chinese AI Safety Evaluation Practice

Independent Project

Jan 2026 - Present (8 months)

Evaluated Chinese-language prompts and model responses using a structured framework covering intent, risk category, adversarial framing, dual-use potential, escalation, severity, decision, and rationale. Classified safety risks and produced concise rationales with recommended outcomes.

TO

Content Review and Teaching Assistant

To Fill Organization

Reviewed educational materials, exercises, answer keys, and explanations for accuracy, clarity, consistency, and formatting quality. Applied predefined standards across high-volume content and documented correction needs.

Education

Degrees, certifications, and relevant coursework

眠眠 hasn't added their education

Don't worry, there are 90k+ talented remote workers on Himalayas

Tech stack

Software and tools used professionally

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan