
AI Safety Practitioner
Posted 6 hours ago

Posted 6 hours ago
This is a fully remote position, open to applicants in United States.
β’ Assess AI-generated responses for safety, factual correctness, adherence to policies, and overall quality.
β’ Analyze content related to misinformation, political influence, self-harm, violence, cyber threats, biosecurity, and other sensitive areas.
β’ Implement and enhance evaluation frameworks for RLHF, SFT, and AI safety benchmarks.
β’ Detect unsafe outputs, hallucinations, logical errors, and breaches of policy.
β’ Offer structured feedback to enhance model alignment and safety performance.
β’ Partner with AI researchers and safety teams on ongoing evaluation projects.
β’ Engage in initiatives that train and improve advanced AI systems.
β’ A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
β’ Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a similar field.
β’ Exceptional proficiency in written English.
β’ Strong critical thinking and analytical reasoning capabilities.
β’ Ability to evaluate nuanced and policy-sensitive scenarios consistently.
β’ Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
β’ Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.
β’ Experience in reviewing complex, high-risk, or ambiguous content is preferred.
β’ Candidates must not require H1-B or STEM OPT sponsorship; Mercor is unable to support H1-B or STEM OPT candidates.
β’ Completely remote position.
β’ Flexible scheduling.
β’ Weekly payments via Stripe or Wise based on services provided.
β’ Project timelines may be extended, shortened, or concluded early based on requirements and performance.
β’ Collaborate with top AI researchers, engineers, and safety teams.
β’ Contribute to shaping the future of AI systems.
β’ Competitive compensation.
β’ Earn up to $400 for each successful referral, with no cap on the number of referrals.
β’ Reasonable accommodations will be made available upon request.
Ensemble Health Partners
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.