
AI Safety Practitioner
Posted Aug 10

Posted Aug 10
This is a fully remote position, open to applicants in United States.
β’ Assess AI-generated outputs for safety, factual correctness, adherence to policies, and overall quality.
β’ Examine content related to misinformation, political influences, self-harm, violence, cyber threats, biosecurity, and other sensitive topics.
β’ Develop and refine evaluation criteria for Reinforcement Learning from Human Feedback (RLHF), Supervised Fine-Tuning (SFT), and benchmarks for AI safety.
β’ Detect unsafe outputs, hallucinations, logical errors, and breaches of policy.
β’ Deliver structured feedback aimed at enhancing model alignment and safety efficacy.
β’ Collaborate with AI researchers and safety teams on ongoing evaluation projects.
β’ Engage in initiatives that train and improve cutting-edge AI systems.
β’ A Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
β’ Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a similar area.
β’ Exceptional written English skills.
β’ Strong critical thinking and analytical reasoning abilities.
β’ Capability to consistently assess nuanced and policy-sensitive situations.
β’ Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
β’ Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.
β’ Experience in reviewing complex, high-risk, or ambiguous content is preferred.
β’ Must not require H1-B or STEM OPT sponsorship.
β’ Fully remote position.
β’ Flexible scheduling.
β’ Project timelines can be adjusted based on needs and performance.
β’ Weekly payments through Stripe or Wise based on services provided.
β’ Competitive compensation.
β’ Work alongside leading AI researchers, engineers, and safety teams.
β’ Opportunity to influence cutting-edge AI models utilized by millions globally.
β’ Up to $400 for each successful referral, with no cap on the number of referrals.
β’ Reasonable accommodations available upon request.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.