
AI Safety Red Teamer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United States.
β’ Create adversarial prompts to rigorously evaluate frontier AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and policy shortcomings.
β’ Assess model resilience in areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive sectors.
β’ Record vulnerabilities and aid in the creation of safety benchmarking and red-teaming documentation.
β’ Partner with AI researchers to enhance model alignment, durability, and safety.
β’ Engage in initiatives aimed at training and refining AI systems.
β’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a similar field.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related area.
β’ Proficient analytical reasoning, prompt design, and written communication abilities.
β’ Experience in crafting adversarial prompts or assessing frontier AI technologies.
β’ Preferred experience in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
β’ Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation techniques is preferred.
β’ Specialized knowledge in one or more grey-area fields, including cybersecurity, biosecurity, political content, misinformation, or scientific safety is preferred.
β’ Must be capable of functioning as an independent contractor.
β’ H1-B and STEM OPT candidates will not be supported.
β’ Fully remote position.
β’ Flexible scheduling.
β’ Weekly payments through Stripe or Wise.
β’ Chance to work on innovative adversarial testing with top AI researchers and safety teams.
β’ Ability to impact how AI systems tackle complex, real-world safety issues.
β’ Project durations may vary, being extended, shortened, or concluded early based on requirements and performance.
β’ Competitive compensation.
β’ Referral bonuses of up to $340 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.