
AI Safety Expert β English, Assamese
Posted Sep 1

Posted Sep 1
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments of conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Create human data by annotating errors, classifying vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Explore sensitive AI topics, including bias, misinformation, and detrimental behaviors.
β’ Detect vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues during production.
β’ Enhance customer AI systems through human data-informed red teaming.
β’ Native proficiency in English and Assamese.
β’ Previous experience in red teaming within AI adversarial contexts, cybersecurity, or socio-technical exploration.
β’ Capability to challenge AI systems adversarially and push them to their limits.
β’ Familiarity with frameworks, taxonomies, benchmarks, or playbooks for systematic testing.
β’ Ability to clearly communicate risks to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects and clients.
β’ Engagement as an independent contractor.
β’ Must not require H1-B or STEM OPT sponsorship.
β’ Fully remote position.
β’ Flexible working hours / ability to work on your own schedule.
β’ Weekly payments via Stripe or Wise.
β’ Optional participation in higher-sensitivity projects.
β’ Clear guidelines and wellness resources available for projects involving sensitive content.
β’ Competitive compensation.
β’ Referral program offering up to $90 for each successful referral.
β’ Reasonable accommodations provided upon request.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.