Remotery

AI Safety Practitioner

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$60 – $70/hour

Posted 6 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Assess AI-generated outputs for safety, factual correctness, adherence to policies, and overall quality.

• Analyze content related to misinformation, political manipulation, self-harm, violence, cybersecurity, biosecurity, and other sensitive areas.

• Develop and refine evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and AI safety benchmarks.

• Detect unsafe outputs, hallucinations, logical errors, and policy breaches.

• Offer structured feedback to enhance model alignment and safety performance.

• Collaborate with AI researchers and safety teams on ongoing evaluation projects.

• Engage in projects aimed at training and improving cutting-edge AI systems.


⛳️ Requirements

• A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.

• Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related discipline.

• Exceptional proficiency in written English.

• Strong critical thinking and analytical reasoning abilities.

• Capability to consistently assess complex and policy-sensitive situations.

• Prior experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation preferred.

• Knowledge of safety policies, content moderation, or the development of evaluation rubrics preferred.

• Background in reviewing complex, high-risk, or ambiguous content preferred.

• Must be able to operate as an independent contractor.

• H1-B and STEM OPT candidates are not eligible.


🏝️ Benefits

• Fully remote work environment.

• Flexibility to set your own schedule.

• Project timelines can be adjusted based on needs and performance, including extensions, reductions, or early conclusions.

• Weekly payments via Stripe or Wise based on services provided.

• Competitive compensation.

• Opportunity to collaborate with prominent AI researchers, engineers, and safety teams.

• Chance to influence frontier AI models utilized by millions globally.

• Up to $400 for each successful referral.

• Unlimited number of referrals allowed.

• Reasonable accommodations available upon request.

People also viewed

Ensemble Health Partners5 hours ago

Engineer II, AI – Copilot Agent

IN flagIndia OnlyFull-timeArtificial Intelligence
ApplyView job
Mercor6 hours ago

AI Safety Practitioner

US flagUnited States OnlyFreelanceArtificial Intelligence$60 – $70/hour
ApplyView job
Mercor6 hours ago

AI Safety Practitioner

US flagUnited States OnlyFreelanceArtificial Intelligence$60 – $70/hour
ApplyView job
Mercor6 hours ago

AI Safety Practitioner

US flagUnited States OnlyFreelanceArtificial Intelligence$60 – $70/hour
ApplyView job
Gartner19 hours ago

Senior Director Analyst – AI and Emerging Technologies Enterprise Strategy

GB flagUnited Kingdom, +4 more statesFull-timeArtificial Intelligence
ApplyView job
Terac21 hours ago

AI Professional – Feedback on Startup Outreach Copy

US flagUnited States OnlyFreelanceArtificial Intelligence$3/hour
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers