
AI Safety Expert β English, Punjabi
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in United States.
β’ Engage in red teaming of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Create high-quality human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Investigate AI systems for vulnerabilities that automated testing may overlook.
β’ Enhance evaluation coverage and minimize unexpected issues in production.
β’ Assist Mercor clients in enhancing the safety, robustness, and reliability of their AI systems.
β’ Proficiency in both English and Punjabi is essential.
β’ Previous experience in red teaming related to AI adversarial work, cybersecurity, or socio-technical assessments is required.
β’ Capability to probe systems in an adversarial manner and challenge them to their limits.
β’ Familiarity with frameworks or benchmarks for systematic testing is necessary.
β’ Ability to clearly communicate risks to both technical and non-technical audiences.
β’ Flexibility to adapt to various projects and client needs.
β’ Must hold independent contractor status.
β’ Currently unable to accommodate H1-B or STEM OPT candidates.
β’ Fully remote position.
β’ Flexibility to create your own schedule.
β’ Weekly payments processed through Stripe or Wise.
β’ Option to participate in higher-sensitivity projects.
β’ Access to clear guidelines and wellness resources for projects involving sensitive content.
β’ Competitive compensation.
β’ Opportunity to collaborate with leading researchers in the field.
β’ Referral bonuses of up to $90 for each successful referral.
CVS Health
One Impression
Volga Partners
Mercor
Get handpicked remote jobs straight to your inbox weekly.