
AI Safety Red Teamer
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in United States.
β’ Create adversarial prompts to evaluate the resilience of advanced AI models.
β’ Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy.
β’ Assess model robustness in areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive sectors.
β’ Record vulnerabilities and assist in the development of safety benchmarking and red-teaming documentation.
β’ Partner with AI researchers to enhance model alignment, robustness, and safety measures.
β’ Engage in projects aimed at training and improving AI systems.
β’ Execute project tasks independently with a flexible schedule.
β’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.
β’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a closely related area.
β’ Proficient analytical reasoning abilities.
β’ Exceptional skills in prompt design.
β’ Strong written communication proficiency.
β’ Experience in crafting adversarial prompts or assessing advanced AI systems.
β’ Preference for candidates with experience in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
β’ Familiarity with jailbreak testing, prompt engineering, or methodologies for adversarial evaluation is preferred.
β’ Expertise in cybersecurity, biosecurity, political content, misinformation, or scientific safety is desirable.
β’ Candidates must not require H1-B or STEM OPT support.
β’ Fully remote position.
β’ Flexible scheduling options.
β’ Weekly compensation through Stripe or Wise based on services provided.
β’ Competitive salary.
β’ Opportunity to collaborate with top-tier researchers.
β’ Engage in groundbreaking AI safety initiatives.
β’ Chance to impact the safety of advanced AI systems.
β’ Reasonable accommodations available upon request.
β’ Referral program offering up to $340 for each successful referral.
CVS Health
One Impression
Volga Partners
Mercor
Get handpicked remote jobs straight to your inbox weekly.