
AI Safety Experts, English β Punjabi
Posted Sep 15

Posted Sep 15
This is a fully remote position, open to applicants in United States.
β’ Conduct red team operations on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.
β’ Document failures, categorize vulnerabilities, and identify systemic risks.
β’ Adhere to established taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Identify vulnerabilities that automated testing may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues in production.
β’ Enhance the safety, robustness, and trustworthiness of customer AI systems.
β’ Fluent or native proficiency in both English and Punjabi.
β’ Previous experience in red teaming focused on AI adversarial work, cybersecurity, or socio-technical probing.
β’ Capability to adversarially probe AI systems and push them to their limits.
β’ Familiarity with frameworks or benchmarks for systematic testing.
β’ Ability to clearly communicate risks to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects and clients.
β’ Status as an independent contractor.
β’ Must not require H1-B or STEM OPT sponsorship.
β’ Fully remote position that allows for flexible scheduling.
β’ Weekly payments via Stripe or Wise based on services provided.
β’ Project timelines can be extended, shortened, or concluded early based on requirements and performance.
β’ Participation in higher-sensitivity projects is optional.
β’ Clear guidelines and wellness resources available for higher-sensitivity projects.
β’ Reasonable accommodations provided upon request.
β’ Competitive compensation.
β’ Referral bonuses of up to $90 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.