Remotery

AI Safety Red Teamer

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$70 – $84/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Create adversarial prompts to rigorously test advanced AI models

β€’ Detect jailbreaks, unsafe behaviors, hallucinations, and failures in policy

β€’ Assess model resilience in areas such as misinformation, cyber threats, biosecurity, fraud, political content, and other sensitive topics

β€’ Record vulnerabilities and assist in creating safety benchmarks and red-teaming documentation

β€’ Partner with AI researchers to enhance model alignment, robustness, and safety protocols

β€’ Engage in projects aimed at training and improving AI systems


⛳️ Requirements

β€’ A Bachelor's degree or higher in fields such as Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related area

β€’ A minimum of 5 years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related discipline

β€’ Excellent analytical reasoning, prompt design, and written communication abilities

β€’ Experience in crafting adversarial prompts or assessing advanced AI systems

β€’ Preferred experience in AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety

β€’ Familiarity with methodologies for jailbreak testing, prompt engineering, or adversarial evaluation is preferred

β€’ Knowledge in one or more grey-area domains, including cyber security, biosecurity, political content, misinformation, or scientific safety is preferred

β€’ Must not require H1-B or STEM OPT support at this time


🏝️ Benefits

β€’ Fully remote work

β€’ Flexible scheduling

β€’ Weekly payments through Stripe or Wise based on services provided

β€’ Opportunity to collaborate with leading AI researchers and safety teams

β€’ Involvement in cutting-edge adversarial testing initiatives

β€’ Opportunity to influence the responses of AI systems to real-world safety challenges

β€’ Competitive compensation

β€’ Referral bonuses of up to $340 for each successful referral

β€’ Reasonable accommodations available upon request

People also viewed

Progressive Leasing7 hours ago

AI Workforce Enablement Consultant – Contract

US flagArizona OnlyFull-timeArtificial Intelligence
ApplyView job
apna9 hours ago

AI Voice Data Collection – Hindi

IN flagIndia OnlyFreelanceArtificial Intelligenceβ‚Ή500/hour
ApplyView job
apna9 hours ago

AI Voice Data Collection – Odia

IN flagIndia OnlyFreelanceArtificial Intelligenceβ‚Ή500/hour
ApplyView job
Texas Research International9 hours ago

Junior Software and Systems Specialist – AI and Automation

US flagUnited States OnlyFull-timeArtificial Intelligence
ApplyView job
Welo Global9 hours ago

Generative AI Analyst, English

GB flagUnited Kingdom OnlyFreelanceArtificial Intelligence$19/hour
ApplyView job
Welo Global9 hours ago

Generative AI Analyst – French

CA flagCanada OnlyFreelanceArtificial Intelligence$20/hour
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers