
AI Safety Practitioner
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United States.
• Assess AI-generated outputs for safety, factual correctness, adherence to policies, and overall quality.
• Examine content related to misinformation, political influences, self-harm, violence, cyber threats, biosecurity, and other sensitive topics.
• Develop and refine evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and benchmarks for AI safety.
• Detect unsafe outputs, hallucinations, logical errors, and breaches of policy.
• Deliver structured feedback aimed at enhancing model alignment and safety efficacy.
• Collaborate with AI researchers and safety teams on ongoing evaluation projects.
• Engage in initiatives that train and improve cutting-edge AI systems.
• A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
• Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a similar area.
• Exceptional written English skills.
• Strong critical thinking and analytical reasoning abilities.
• Capability to consistently assess nuanced and policy-sensitive situations.
• Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
• Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.
• Experience in reviewing complex, high-risk, or ambiguous content is preferred.
• Must not require H1-B or STEM OPT sponsorship.
• Fully remote position.
• Flexible scheduling.
• Project timelines can be adjusted based on needs and performance.
• Weekly payments through Stripe or Wise based on services provided.
• Competitive compensation.
• Work alongside leading AI researchers, engineers, and safety teams.
• Opportunity to influence cutting-edge AI models utilized by millions globally.
• Up to $400 for each successful referral, with no cap on the number of referrals.
• Reasonable accommodations available upon request.
Mercor
Cloudera
NextLink Group
Get handpicked remote jobs straight to your inbox weekly.