
AI Safety Red Teamer
Posted Aug 31

Posted Aug 31
This is a fully remote position, open to applicants in United States.
β’ Create adversarial prompts to thoroughly assess cutting-edge AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy adherence.
β’ Assess model resilience in areas such as misinformation, cyber threats, biosecurity, fraud, political content, and other sensitive subjects.
β’ Record vulnerabilities and assist in the development of safety benchmarks and red-teaming documentation.
β’ Work collaboratively with AI researchers to enhance model alignment, robustness, and safety.
β’ Engage in projects aimed at training and refining AI systems.
β’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a similar area.
β’ Excellent analytical reasoning, prompt design, and written communication abilities.
β’ Experience in crafting adversarial prompts or assessing cutting-edge AI systems.
β’ Preferred: experience in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety practices.
β’ Preferred: knowledge of jailbreak testing, prompt engineering, or adversarial evaluation techniques.
β’ Preferred: proficiency in one or more grey-area fields, including cyber threats, biosecurity, political content, misinformation, or scientific safety.
β’ Ability to work as an independent contractor is essential.
β’ H1-B and STEM OPT candidates cannot be accommodated.
β’ Completely remote position.
β’ Flexible scheduling.
β’ Weekly compensation via Stripe or Wise.
β’ Project duration may be extended, shortened, or concluded early based on requirements and performance.
β’ Competitive salary.
β’ Collaborate with top AI researchers and safety teams.
β’ Opportunity to engage in advanced adversarial testing.
β’ Chance to impact AI safety and model evolution.
β’ Referral bonus of up to $340 for each successful referral.
β’ Reasonable accommodations provided upon request.
24-MAG
Keppri
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.