AI Safety Data Reviewer Japanese English
About OpenTrain
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover meaningful projects that shape how modern AI systems work.
Creating an OpenTrain account is free, and candidates can apply in minutes for remote opportunities across the growing AI training industry.
About AI Safety Training
AI training is the human side of building artificial intelligence. Specialists review model responses, test difficult scenarios, and provide structured feedback so AI systems become more accurate, useful, and safe.
In this role, your judgments will help identify unsafe, adversarial, toxic, or misleading outputs before models are used in real-world settings.
The Role
OpenTrain AI is seeking an AI Safety Data Reviewer with near-native or native Japanese and strong English skills. This remote, hourly-paid contractor role focuses on evaluating AI-generated content, safety decisions, reasoning quality, and step-by-step problem-solving across Japanese and English content.
You will make nuanced judgments about policy alignment, cultural context, and response quality. The work is part-time at 20 or more hours per week and pays $27-$31 per hour, with an hourly rate of $30.
- Contractor and part-time engagement
- Remote work available worldwide
- 20+ hours per week
- $27-$31 per hour
What You'll Do
You will assess multiple AI responses and provide clear, reproducible feedback that helps improve model safety and performance. Reviews may involve Japanese, English, or content requiring careful comparison across both languages.
- Review AI-generated content and safety decisions
- Evaluate solutions for correctness, clarity, and logical reasoning
- Assess step-by-step problem-solving and identify methodological or conceptual errors
- Fact-check responses as needed
- Rate or compare multiple responses for safety and policy alignment
- Identify edge cases, adversarial behavior, and potential safety failures
- Account for cultural nuance, slang, coded language, and context shifts
- Recommend mitigations for unsafe or problematic model behavior
- Write clear rationales for moderation and safety decisions
Requirements
This role requires advanced bilingual judgment, strong analytical writing, and substantial experience in trust and safety or a related safety function. A bachelor's degree or higher in a relevant field is expected, unless you have equivalent professional experience.
- Near-native or native Japanese proficiency in reading and writing
- Minimum C1 English proficiency in reading and writing
- Bachelor's degree or higher in Communications, Linguistics, Psychology, Law or Policy, Security Studies, or a related field, or equivalent professional experience
- Senior-level experience in Trust & Safety, content moderation, policy operations, risk, compliance, investigations, or related safety work
- Proven LLM red-teaming or adversarial testing experience
- Strong knowledge of hate and harassment, sexual content, self-harm, violence, bias, illegal goods or services, malicious activities, malicious code, and misinformation
- Experience applying policy standards consistently across Japanese and English content
- Strong analytical writing with clear, reproducible decision rationales
Content Exposure and Working Environment
This work may expose you to explicit, toxic, violent, sexual, or psychologically disturbing material. You should be comfortable reviewing potentially disturbing content, including sexual or violent topics, in a secure remote work environment while applying safety standards consistently.
- Potential exposure to adversarial, toxic, sexual, violent, or unsafe content
- Secure remote work environment required
- Careful, consistent judgment across sensitive policy areas
Preferred Experience
Localization or translation experience is preferred. The ability to preserve meaning, severity, and intent across languages is especially valuable when reviewing safety-sensitive content.
- Japanese and English content review
- Localization or translation experience
- Understanding of cultural nuance, slang, coded language, and context shifts
- Experience identifying edge cases and recommending mitigations
How to Apply
Create a free OpenTrain account and apply through OpenTrain AI. If selected, you will contribute directly to the human feedback and evaluation work that helps make advanced AI systems safer and more dependable.
- Apply remotely through OpenTrain
- Work part time with a 20+ hour weekly commitment
- Use bilingual safety expertise to influence AI behavior