AI Safety Red Teamer

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$70 – $84/hour

Posted 21 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Develop adversarial prompts to rigorously test cutting-edge AI models

• Detect jailbreaks, unsafe behaviors, hallucinations, and policy violations

• Assess model resilience in areas such as misinformation, cybersecurity, biosecurity, fraud, political content, and various sensitive sectors

• Record vulnerabilities and contribute to safety benchmarking and red-team analysis reports

• Partner with AI researchers to enhance model alignment, resilience, and safety

• Engage in projects aimed at training and refining AI systems


⛳️ Requirements

• Bachelor’s degree or higher in fields such as Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related area

• Over 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a comparable field

• Excellent analytical reasoning, prompt design, and written communication skills

• Proven experience in designing adversarial prompts or assessing cutting-edge AI systems

• Preferred: background in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety

• Preferred: knowledge of jailbreak testing, prompt engineering, or adversarial evaluation techniques

• Preferred: specialization in grey-area domains such as cybersecurity, biosecurity, political content, misinformation, or scientific safety

• Must be capable of functioning as an independent contractor

• H1-B and STEM OPT candidates cannot be accommodated


🏝️ Benefits

• Fully remote position

• Flexible schedule / can be managed according to your own timetable

• Weekly payments through Stripe or Wise based on services provided

• Projects may be extended, shortened, or concluded early based on requirements and performance

• Competitive compensation

• Chance to collaborate with leading AI researchers and safety teams

• Opportunity to impact the advancement of cutting-edge AI systems

• Referral program offering up to $340 for each successful referral

People also viewed

Mercor13 hours ago

Software, AI, IT, Data Evaluator

US flagUnited States OnlyFreelanceArtificial Intelligence$80 – $120/hour
ApplyView job
Vehlo16 hours ago

Director, Information Technology – Enterprise AI

US flagFlorida, +4 more statesFull-timeArtificial Intelligence
ApplyView job
Lifebit17 hours ago

General Manager – AI Business Unit

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job
Gartner23 hours ago

Senior Director Analyst – Integration and APIs, AI

US flagUnited States OnlyFull-timeArtificial Intelligence$172k – $202.5k/year
ApplyView job
Gartner23 hours ago

Senior Director, Analyst – Healthcare Providers, RCM & Healthcare AI Technology

US flagUnited States OnlyFull-timeArtificial Intelligence$172k – $202.5k/year
ApplyView job
Gartner23 hours ago

Senior Director, Analyst – Tech CEO Business and Strategy Advisor on AI Orchestration

US flagTexas OnlyFull-timeArtificial Intelligence$172k – $202.5k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers