AI Safety Red Teamer

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$70 – $84/hour

Posted 8 hours ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Create adversarial prompts to evaluate and challenge cutting-edge AI models.

β€’ Detect jailbreaks, unsafe behaviors, hallucinations, and policy shortcomings.

β€’ Assess the robustness of models across misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive areas.

β€’ Record vulnerabilities and aid in safety benchmarking and red-teaming documentation.

β€’ Work in collaboration with AI researchers to enhance model alignment, robustness, and safety.

β€’ Engage in projects aimed at training and improving AI systems.

β€’ Undertake independent-contractor assignments on a flexible schedule; project durations may vary based on requirements and performance.


⛳️ Requirements

β€’ A Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field.

β€’ Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a comparable area.

β€’ Excellent analytical reasoning, prompt design, and written communication abilities.

β€’ Proven experience in designing adversarial prompts or assessing frontier AI systems.

β€’ Preferred experience in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.

β€’ Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation techniques is preferred.

β€’ Specialized knowledge in one or more grey-area domains, including cybersecurity, biosecurity, political content, misinformation, or scientific safety is preferred.

β€’ Must be eligible to work in one of the specified locations.

β€’ H1-B and STEM OPT candidates are not eligible.


🏝️ Benefits

β€’ Fully remote position that allows you to work on your own schedule.

β€’ Weekly payments through Stripe or Wise based on services provided.

β€’ Competitive compensation.

β€’ Collaboration with prominent researchers.

β€’ Chance to contribute to the development of the next generation of AI systems.

β€’ Referral bonuses of up to $340 for each successful referral.

People also viewed

GE Vernova7 hours ago

AI Portfolio and Governance Leader

US flagUnited States OnlyFull-timeArtificial Intelligence$114.1k – $190.2k/year
ApplyView job
ICS AI Ltd8 hours ago

AI Service Analyst

GB flagUnited Kingdom OnlyFull-timeArtificial IntelligenceΒ£26.5k – Β£30k/year
ApplyView job
Mercor8 hours ago

AI Safety Expert, English, Malay

US flagUnited States OnlyFreelanceArtificial Intelligence$17 – $25/hour
ApplyView job
Mercor8 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
Mercor8 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
The Hershey Company9 hours ago

Enterprise AI Control Tower Owner

US flagPennsylvania OnlyFull-timeArtificial Intelligence$114.6k – $143.3k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers