About this role
Role Overview
Help strengthen the safety of advanced AI systems by applying Japanese language fluency and cultural judgment to how models respond to sensitive topics. Training on the required workflow is provided, and prior AI or machine learning experience is not required.
Key Responsibilities
- Evaluate and improve AI model handling of sensitive topics in Japanese.
- Write expert level Japanese prompts across a variety of sensitive subject areas.
- Use structured guidelines to classify prompts and conversations.
- Identify adversarial phrasing and escalation patterns, then document the reasoning behind each judgment.
Qualifications
- Native or near native Japanese fluency and business level written English.
- Bachelor''s degree completed or currently in progress.
- Strong written reasoning skills and close attention to detail.
- Sound judgment when working with sensitive and dual use information.
Preferred Qualifications
- Based in Japan or the broader East Asia region, though applicants elsewhere are welcome.
- Experience reviewing, grading, or red teaming written or technical content.
- Background in trust and safety, content moderation, policy evaluation, or adversarial testing.
Work Terms
- Remote, with East Asia preferred.
- Hourly engagement.
- Immediate start.
Compensation
$48 to $52 per hour.