
AI Safety Red Teamer
Posted 14 hours ago

Posted 14 hours ago
This is a fully remote position, open to applicants in United States.
• Create adversarial prompts to challenge and assess frontier AI models.
• Detect jailbreaks, unsafe behaviors, hallucinations, and policy shortcomings.
• Analyze model resilience in areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive topics.
• Record vulnerabilities and assist in safety benchmarking and red-teaming documentation.
• Collaborate with AI researchers to enhance model alignment, robustness, and safety measures.
• Manage project timelines independently as a contractor.
• Engage in projects aimed at training and improving AI systems.
• A Bachelor’s degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
• Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related area.
• Excellent analytical reasoning, prompt design skills, and written communication abilities.
• Proven experience in crafting adversarial prompts or evaluating advanced AI systems.
• Must be capable of functioning as an independent contractor.
• H1-B and STEM OPT candidates are not eligible.
• Fully remote work environment.
• Flexible scheduling options.
• Weekly payments through Stripe or Wise.
• Chance to collaborate with leading AI researchers and safety teams.
• Work on innovative AI safety initiatives.
• Opportunity to impact the safety of frontier AI models.
• Referral bonus of up to $340 for each successful referral.
Mercor
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.