
AI Safety Practitioner
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in United States.
• Assess AI-generated outputs for safety, accuracy, adherence to policies, and overall quality.
• Examine content related to misinformation, political manipulation, self-harm, violence, cybersecurity, biosecurity, and other sensitive topics.
• Implement and enhance evaluation standards for RLHF, SFT, and AI safety metrics.
• Detect unsafe outputs, hallucinations, logical errors, and breaches of policy.
• Deliver structured feedback aimed at improving model alignment and safety efficacy.
• Collaborate with AI researchers and safety teams on ongoing evaluation projects.
• Engage in initiatives focused on training and improving AI systems.
• A Bachelor’s degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
• Over 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a similar area.
• Exceptional written English skills.
• Strong critical thinking and analytical reasoning abilities.
• Capability to consistently assess nuanced and policy-sensitive situations.
• Preferred experience in AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
• Familiarity with safety policies, content moderation, or the development of evaluation rubrics is preferred.
• Preferred experience in reviewing complex, high-risk, or ambiguous content.
• Must not require H1-B or STEM OPT support.
• Fully remote position.
• Flexible schedule; work according to your own timetable.
• Weekly payments via Stripe or Wise based on services rendered.
• Projects may be extended, shortened, or concluded early based on needs and performance.
• Collaborate with top AI researchers, engineers, and safety teams.
• Opportunity to influence the safety and behavior of cutting-edge AI models.
• Competitive compensation.
• Earn up to $400 for each successful referral, with no cap on the number of referrals.
• Reasonable accommodations available upon request.
Progressive Leasing
apna
apna
Texas Research International
Get handpicked remote jobs straight to your inbox weekly.