
AI Safety Experts, English β Telugu
Posted 5 days ago

Posted 5 days ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.
β’ Create human-centric data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.
β’ Generate reproducible reports, datasets, and attack case studies.
β’ Review AI outputs that involve sensitive subjects such as bias, misinformation, or harmful behaviors.
β’ Detect vulnerabilities that automated tests may overlook.
β’ Enhance evaluation coverage and minimize unexpected issues in production.
β’ Collaborate on projects aimed at training and improving AI systems.
β’ Proficiency in English and Telugu is essential.
β’ Strong discernment regarding language and content.
β’ Capability to evaluate whether AI responses are accurate, thorough, and suitable.
β’ Skill in identifying subtle errors, inconsistencies, and omissions.
β’ Consistent adherence to guidelines and quality standards.
β’ Ability to articulate reasoning clearly to both technical and non-technical stakeholders.
β’ Flexibility across varying projects, task types, and clientele.
β’ Status as an independent contractor.
β’ Not eligible as an H1-B or STEM OPT candidate.
β’ Preferred expertise in adversarial ML, cybersecurity, socio-technical risk, or creative probing.
β’ Fully remote position.
β’ Flexible working hours allowing you to set your own schedule.
β’ Weekly payments through Stripe or Wise.
β’ Competitive compensation.
β’ Access to wellness resources and clear protocols for higher-sensitivity projects.
β’ Reasonable accommodations available upon request.
β’ Referral bonuses of up to $90 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.