
AI Safety Red Teamer
Posted 14 hours ago

Posted 14 hours ago
This is a fully remote position, open to applicants in United States.
β’ Create adversarial prompts to challenge and evaluate frontier AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy compliance.
β’ Assess model resilience in the contexts of misinformation, cybersecurity, biosecurity, fraud, political content, and other critical areas.
β’ Record vulnerabilities and assist in the development of safety benchmarking and red-teaming reports.
β’ Collaborate with AI researchers to enhance model alignment, robustness, and safety measures.
β’ Engage in projects focused on training and improving frontier AI systems.
β’ A Bachelor's degree or higher in fields such as Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a similar area.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related domain.
β’ Exceptional skills in analytical reasoning, prompt design, and written communication.
β’ Proven experience in designing adversarial prompts or assessing frontier AI systems.
β’ Required to be based in one of the specified eligible countries or territories.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Completely remote work environment.
β’ Flexible working hours, allowing you to set your own schedule.
β’ Weekly payments through Stripe or Wise.
β’ Opportunities for project extensions.
β’ Competitive compensation.
β’ Collaboration with top AI researchers and safety teams.
β’ Reasonable accommodations available upon request.
β’ Referral bonuses of up to $340 for each successful referral.
Mercor
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.