
AI Safety Red Teamer
Posted 4 days ago

Posted 4 days ago
This is a fully remote position, open to applicants in United States.
β’ Develop adversarial prompts to thoroughly evaluate cutting-edge AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and policy breaches.
β’ Assess model resilience in the context of misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive areas.
β’ Record vulnerabilities and assist in creating safety benchmarking and red-teaming documentation.
β’ Collaborate with AI researchers to enhance model alignment, robustness, and safety.
β’ Engage in initiatives aimed at training and refining AI systems.
β’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related area.
β’ Exceptional analytical reasoning, prompt design, and written communication abilities.
β’ Experience in crafting adversarial prompts or assessing advanced AI systems.
β’ Currently, we are unable to accommodate H1-B or STEM OPT candidates.
β’ Fully remote work opportunities.
β’ Flexible scheduling options.
β’ Weekly payments through Stripe or Wise.
β’ Competitive compensation.
β’ Project timelines may be adjusted based on needs and performance.
β’ Referral program: earn up to $340 for each successful referral.
XenoPatch GmbH
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.