About this role
Help strengthen frontier AI systems by uncovering vulnerabilities through rigorous adversarial testing. In this role, you will develop challenging prompts, expose model weaknesses, and assess AI behavior in complex, high-risk, and ambiguous areas.
Key Responsibilities- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behavior, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarks and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety.
- Bachelor''s degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI safety, AI red teaming, trust and safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems.
- Experience with AI red teaming, RLHF, SFT, AI alignment, or trust and safety.
- Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methods.
- Expertise in one or more ambiguous or high-risk domains, including cybersecurity, biosecurity, political content, misinformation, or scientific safety.
- Remote hourly engagement.
- $70 to $84 per hour.
- Contribute to securing the next generation of frontier AI models alongside AI researchers and safety teams.
- Help shape how AI systems handle complex, real-world safety challenges.