
AI Safety Expert β English, Punjabi
Posted 4 days ago

Posted 4 days ago
This is a fully remote position, open to applicants in United States.
β’ Engage in red-teaming of conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Document failures, categorize vulnerabilities, and identify systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to maintain consistent testing protocols.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Review AI outputs related to sensitive subjects like bias, misinformation, or harmful behaviors.
β’ Assist in broadening evaluation coverage and minimizing surprises during production.
β’ Participate in projects aimed at training and enhancing advanced AI systems for Mercorβs clientele.
β’ Must have native fluency in both English and Punjabi.
β’ Demonstrate strong judgment regarding language and content.
β’ Capable of evaluating whether AI responses are accurate, complete, and suitable, along with providing explanations.
β’ Skilled in detecting subtle errors, inconsistencies, and gaps.
β’ Consistently adhere to guidelines and quality standards.
β’ Ability to communicate reasoning clearly to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, task types, and client needs.
β’ Must be available as an independent contractor.
β’ Candidates must not be on H1-B or STEM OPT visas.
β’ Preferred experience in adversarial machine learning, cybersecurity, socio-technical risks, or creative probing.
β’ Fully remote position.
β’ Flexibility to work on your own schedule.
β’ Receive weekly payments via Stripe or Wise based on the services performed.
β’ Optional involvement in projects with higher sensitivity.
β’ Clear guidelines and wellness resources available for higher-sensitivity projects.
β’ Reasonable accommodations provided upon request.
β’ Earn referral payments of up to $90 for each successful referral, subject to certain limits.
β’ Opportunity to collaborate with top researchers in the field.
β’ Gain experience in human data-driven AI red teaming at the forefront of safety.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.