
AI Safety Practitioner
Posted 6 hours ago

Posted 6 hours ago
This is a fully remote position, open to applicants in United States.
• Assess AI-generated outputs for safety, factual correctness, adherence to policies, and overall quality.
• Analyze content related to misinformation, political manipulation, self-harm, violence, cybersecurity, biosecurity, and other sensitive areas.
• Develop and refine evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and AI safety benchmarks.
• Detect unsafe outputs, hallucinations, logical errors, and policy breaches.
• Offer structured feedback to enhance model alignment and safety performance.
• Collaborate with AI researchers and safety teams on ongoing evaluation projects.
• Engage in projects aimed at training and improving cutting-edge AI systems.
• A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
• Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related discipline.
• Exceptional proficiency in written English.
• Strong critical thinking and analytical reasoning abilities.
• Capability to consistently assess complex and policy-sensitive situations.
• Prior experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation preferred.
• Knowledge of safety policies, content moderation, or the development of evaluation rubrics preferred.
• Background in reviewing complex, high-risk, or ambiguous content preferred.
• Must be able to operate as an independent contractor.
• H1-B and STEM OPT candidates are not eligible.
• Fully remote work environment.
• Flexibility to set your own schedule.
• Project timelines can be adjusted based on needs and performance, including extensions, reductions, or early conclusions.
• Weekly payments via Stripe or Wise based on services provided.
• Competitive compensation.
• Opportunity to collaborate with prominent AI researchers, engineers, and safety teams.
• Chance to influence frontier AI models utilized by millions globally.
• Up to $400 for each successful referral.
• Unlimited number of referrals allowed.
• Reasonable accommodations available upon request.
Ensemble Health Partners
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.