AI Safety Red Teamer

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$70 – $84/hour

Posted 6 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Develop adversarial prompts to rigorously evaluate cutting-edge AI models

β€’ Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy adherence

β€’ Assess model resilience in areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and other sensitive sectors

β€’ Record vulnerabilities and assist in compiling safety benchmarking and red-teaming documentation

β€’ Partner with AI researchers to enhance model alignment, robustness, and safety

β€’ Execute project tasks as an independent contractor

β€’ Engage in projects aimed at training and refining AI systems


⛳️ Requirements

β€’ A bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related field

β€’ At least 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a similar area

β€’ Excellent analytical reasoning, prompt design, and written communication abilities

β€’ Proven experience in designing adversarial prompts or assessing advanced AI systems

β€’ Preferred: Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety

β€’ Preferred: Understanding of jailbreak testing, prompt engineering, or adversarial evaluation techniques

β€’ Preferred: Specialized knowledge in one or more grey-area domains, such as cybersecurity, biosecurity, political content, misinformation, or scientific safety

β€’ Must not require H1-B or STEM OPT sponsorship/support


🏝️ Benefits

β€’ Fully remote position

β€’ Flexible work schedule / ability to manage your own time

β€’ Weekly payments through Stripe or Wise

β€’ Chance to collaborate with leading AI researchers and safety teams

β€’ Opportunity to impact frontier AI safety and alignment

β€’ Competitive compensation

β€’ Referral bonuses of up to $340 for each successful referral

People also viewed

Mercor20 hours ago

Software, AI, IT, Data Evaluator

US flagUnited States OnlyFreelanceArtificial Intelligence$80 – $120/hour
ApplyView job
Riva Scientific1 day ago

AI Operations Engineer

US flagUnited States OnlyFull-timeArtificial Intelligence
ApplyView job
Mercor1 day ago

AI Safety Expert, English – Punjabi

US flagUnited States OnlyFreelanceArtificial Intelligence$20 – $22/hour
ApplyView job
Mercor1 day ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$20 – $22/hour
ApplyView job
Mercor1 day ago

AI Safety Expert – English, Punjabi

US flagUnited States OnlyFreelanceArtificial Intelligence$20 – $22/hour
ApplyView job
Sigma AI1 day ago

Linguistic Projects – Hindi, Tamil, English

IN flagIndia OnlyFreelanceArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers