
AI Safety Practitioner
Posted 6 hours ago

Posted 6 hours ago
This is a fully remote position, open to applicants in United States.
• Assess AI-generated outputs for safety, accuracy, adherence to policies, and overall quality.
• Examine content related to misinformation, political influence, self-harm, violence, cyber threats, biosecurity, and other sensitive areas.
• Implement and enhance evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and AI safety benchmarking.
• Detect unsafe outputs, hallucinations, logical failures, and breaches of policy.
• Offer structured feedback aimed at enhancing model alignment and safety performance.
• Collaborate with AI researchers and safety teams on ongoing evaluation projects.
• Participate in initiatives focused on training and improving advanced AI systems.
• A Bachelor's degree or higher in fields such as Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related area.
• Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a comparable field.
• Exceptional command of written English.
• Strong critical thinking and analytical reasoning abilities.
• Capability to consistently evaluate complex and policy-sensitive situations.
• Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
• Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.
• Previous experience in reviewing intricate, high-risk, or ambiguous content is preferred.
• Must be an independent contractor.
• H1-B and STEM OPT candidates are not eligible.
• Fully remote position.
• Flexible scheduling.
• Weekly payments through Stripe or Wise based on services rendered.
• Competitive compensation.
• Collaboration with top AI researchers, engineers, and safety teams.
• Opportunity to influence cutting-edge AI models utilized by millions globally.
• Project durations may be extended, shortened, or concluded early based on requirements and performance.
• Referral bonus of up to $400 for each successful referral, with no cap on the number of referrals.
• Reasonable accommodations available upon request.
Ensemble Health Partners
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.