AI Safety Practitioner

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$60 – $70/hour

Posted Aug 10

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Assess AI-generated outputs for safety, factual correctness, adherence to policies, and overall quality.

β€’ Examine content related to misinformation, political influences, self-harm, violence, cyber threats, biosecurity, and other sensitive topics.

β€’ Develop and refine evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and benchmarks for AI safety.

β€’ Detect unsafe outputs, hallucinations, logical errors, and breaches of policy.

β€’ Deliver structured feedback aimed at enhancing model alignment and safety efficacy.

β€’ Collaborate with AI researchers and safety teams on ongoing evaluation projects.

β€’ Engage in initiatives that train and improve cutting-edge AI systems.


⛳️ Requirements

β€’ A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.

β€’ Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a similar area.

β€’ Exceptional written English skills.

β€’ Strong critical thinking and analytical reasoning abilities.

β€’ Capability to consistently assess nuanced and policy-sensitive situations.

β€’ Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.

β€’ Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.

β€’ Experience in reviewing complex, high-risk, or ambiguous content is preferred.

β€’ Must not require H1-B or STEM OPT sponsorship.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible scheduling.

β€’ Project timelines can be adjusted based on needs and performance.

β€’ Weekly payments through Stripe or Wise based on services provided.

β€’ Competitive compensation.

β€’ Work alongside leading AI researchers, engineers, and safety teams.

β€’ Opportunity to influence cutting-edge AI models utilized by millions globally.

β€’ Up to $400 for each successful referral, with no cap on the number of referrals.

β€’ Reasonable accommodations available upon request.

People also viewed

WON.ai21 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board22 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor23 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor23 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor23 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner23 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers