
AI Safety Expert β English, Tamil
Posted Sep 19

Posted Sep 19
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team testing on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Evaluate AI outputs related to sensitive subjects such as bias, misinformation, and harmful behaviors.
β’ Detect vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and mitigate unexpected production issues.
β’ Enhance customer AI systems through adversarial testing.
β’ Native proficiency in both English and Tamil is essential.
β’ Strong discernment regarding language and content; capability to evaluate whether AI responses are precise, comprehensive, and suitable.
β’ Ability to articulate reasoning clearly to both technical and non-technical audiences.
β’ Skill in identifying subtle errors, inconsistencies, and omissions.
β’ Capacity to consistently adhere to guidelines and quality standards.
β’ Flexibility to adapt across various projects, task types, and client needs.
β’ Engagement as an independent contractor.
β’ Currently unable to accommodate H1-B or STEM OPT candidates.
β’ Fully remote position that allows you to work on your own schedule.
β’ Weekly payments via Stripe or Wise based on services provided.
β’ Participation in higher-sensitivity projects is optional.
β’ Clear guidelines and wellness resources available for projects involving sensitive content.
β’ Reasonable accommodations can be arranged upon request.
β’ Competitive compensation.
β’ Referral bonuses of up to $90 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.