
AI Safety Red Teamer
Posted Sep 7

Posted Sep 7
This is a fully remote position, open to applicants in United States.
β’ Develop adversarial prompts to rigorously test cutting-edge AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and policy breaches.
β’ Assess model resilience across areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive sectors.
β’ Record vulnerabilities and assist in creating safety benchmarking and red-teaming documentation.
β’ Partner with AI researchers to enhance model alignment, resilience, and safety.
β’ Engage in projects aimed at training and refining AI systems.
β’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a similar domain.
β’ Excellent analytical reasoning, prompt design, and written communication skills.
β’ Experience in crafting adversarial prompts or assessing cutting-edge AI systems.
β’ Must operate as an independent contractor.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Fully remote position.
β’ Flexibility to work on your own schedule.
β’ Weekly payments via Stripe or Wise based on services rendered.
β’ Project durations may be extended, shortened, or concluded early according to needs and performance.
β’ Competitive compensation.
β’ Collaborate with top researchers and safety teams.
β’ Chance to impact next-generation AI systems.
β’ Up to $340 for each successful referral, with no cap on the number of referrals (restrictions may apply).
β’ Reasonable accommodations available upon request.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.