
AI Safety Expert, English, Portuguese
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red team assessments on conversational AI models and agents using techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Generate human data by annotating failures, identifying vulnerabilities, and flagging systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.
β’ Create reproducible reports, datasets, and attack scenarios for clients.
β’ Evaluate AI outputs related to sensitive issues such as bias, misinformation, or harmful behaviors.
β’ Identify vulnerabilities that automated testing may overlook.
β’ Increase evaluation coverage and minimize unexpected issues in production.
β’ Enhance customer AI systems through adversarial testing.
β’ Native proficiency in English and Portuguese (global, excluding Brazilian Portuguese).
β’ Previous experience with red teaming in AI adversarial contexts, cybersecurity, or socio-technical probing.
β’ Capability to challenge AI systems adversarially and push them to their limits.
β’ Proficiency in using frameworks or benchmarks for structured testing methodologies.
β’ Skill in articulating risks clearly to both technical and non-technical audiences.
β’ Flexibility to adapt to various projects and client needs.
β’ Independent contractor status required.
β’ Must not need H1-B or STEM OPT sponsorship.
β’ Nice-to-have expertise: adversarial machine learning, cybersecurity, socio-technical risk analysis, or innovative probing techniques.
β’ Fully remote position.
β’ Flexible work schedule allowing you to manage your own hours.
β’ Weekly payments processed via Stripe or Wise.
β’ Project timelines may be adjusted, shortened, or concluded early based on requirements and performance.
β’ Participation in higher-sensitivity projects is optional.
β’ Clear guidelines and wellness resources available for projects involving sensitive content.
β’ Competitive salary.
β’ Opportunity to gain experience in human data-driven AI red teaming.
β’ Collaborate with top researchers in the field.
β’ Reasonable accommodations can be made upon request.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.