
AI Safety Expert, English, Odia
Posted Sep 19

Posted Sep 19
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments of conversational AI models and agents by utilizing jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
β’ Create human data through the annotation of failures, classification of vulnerabilities, and identification of systemic risks.
β’ Implement taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Evaluate AI outputs related to sensitive topics such as bias, misinformation, or harmful behaviors.
β’ Identify vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues in production.
β’ Enhance customer AI systems through adversarial testing.
β’ Native fluency in English and Odia.
β’ Strong discernment regarding language and content.
β’ Capability to evaluate if AI responses are accurate, complete, and suitable, along with the ability to articulate the reasoning behind the assessment.
β’ Meticulous attention to detail in identifying subtle errors, inconsistencies, and omissions.
β’ Consistent adherence to guidelines and quality standards.
β’ Skill in clearly communicating reasoning to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, tasks, and client needs.
β’ Status as an independent contractor.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Preferred specializations include adversarial machine learning, cybersecurity, socio-technical risk, and creative probing.
β’ Work fully remote.
β’ Enjoy a flexible schedule and the ability to work at your own pace.
β’ Receive weekly payments through Stripe or Wise.
β’ Project durations may be extended, shortened, or concluded early based on needs and performance.
β’ Access wellness resources and clear guidelines for projects with higher sensitivity.
β’ Earn referral payments of up to $90 for each successful referral.
β’ Reasonable accommodations available upon request.
β’ Collaborate with leading researchers in the field.
β’ Engage in groundbreaking AI safety projects.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.