
AI Safety Red Teamer
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in United States.
• Create adversarial prompts to evaluate the resilience of cutting-edge AI models.
• Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy implementation.
• Assess model robustness in relation to misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive areas.
• Record vulnerabilities and assist in compiling safety benchmarking and red-teaming documentation.
• Work collaboratively with AI researchers to enhance model alignment, robustness, and safety measures.
• Engage in projects aimed at training and improving AI systems.
• A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
• Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a similar area.
• Excellent analytical reasoning, prompt design capabilities, and written communication skills.
• Proven experience in designing adversarial prompts or evaluating advanced AI systems.
• Engagement as an independent contractor.
• Currently unable to support H1-B or STEM OPT candidates.
• Fully remote position.
• Flexible scheduling options.
• Project timelines may be adjusted based on requirements and performance.
• Weekly payments through Stripe or Wise for services provided.
• Competitive compensation.
• Collaborate with top researchers in the field.
• Opportunity to influence the development of future AI systems.
• Reasonable accommodations available upon request.
CVS Health
One Impression
Volga Partners
Mercor
Get handpicked remote jobs straight to your inbox weekly.