AI Safety Red Teamer

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$70 – $84/hour

Posted 8 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Create adversarial prompts to rigorously assess advanced AI models

• Detect jailbreaks, unsafe behaviors, hallucinations, and policy shortcomings

• Assess model resilience across sensitive areas such as misinformation, cyber threats, biosecurity, fraud, and political content

• Record vulnerabilities and assist in the development of safety benchmarks and red-teaming reports

• Collaborate with AI researchers to enhance model alignment, robustness, and safety

• Engage in projects aimed at training and improving AI systems


⛳️ Requirements

• A Bachelor’s degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field

• Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related area

• Strong skills in analytical reasoning, prompt design, and written communication

• Proven experience in designing adversarial prompts or evaluating state-of-the-art AI systems

• Background in AI Red Teaming, Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), AI Alignment, or Trust & Safety

• Knowledge of jailbreak testing, prompt engineering, or adversarial evaluation methods

• Proficiency in one or more grey-area domains, such as cyber security, biosecurity, political content, misinformation, or scientific safety

• Ability to work as an independent contractor

• H1-B and STEM OPT candidates are not eligible


🏝️ Benefits

• Fully remote position that allows you to work on your own schedule

• Weekly payments through Stripe or Wise based on services provided

• Flexibility to extend, shorten, or conclude projects early based on requirements and performance

• Chance to engage in cutting-edge adversarial testing with top AI researchers and safety teams

• Opportunity to shape how AI systems address complex, real-world safety dilemmas

• Competitive compensation

• Referral bonuses of up to $340 for each successful referral

People also viewed

GE Vernova7 hours ago

AI Portfolio and Governance Leader

US flagUnited States OnlyFull-timeArtificial Intelligence$114.1k – $190.2k/year
ApplyView job
ICS AI Ltd8 hours ago

AI Service Analyst

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence£26.5k – £30k/year
ApplyView job
Mercor8 hours ago

AI Safety Expert, English, Malay

US flagUnited States OnlyFreelanceArtificial Intelligence$17 – $25/hour
ApplyView job
Mercor8 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
Mercor8 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
The Hershey Company9 hours ago

Enterprise AI Control Tower Owner

US flagPennsylvania OnlyFull-timeArtificial Intelligence$114.6k – $143.3k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers