
AI Safety Expert β English, Malay
Posted Sep 15

Posted Sep 15
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments of conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Create human data by annotating deficiencies, categorizing vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Evaluate AI outputs related to sensitive subjects, including bias, misinformation, and harmful behaviors.
β’ Assist in broadening evaluation coverage and identifying vulnerabilities that automated tests may overlook.
β’ Collaborate with top researchers and play a role in training and improving advanced AI systems.
β’ Required fluent/native proficiency in both English and Malay.
β’ Previous experience in red teaming, particularly in AI adversarial work, cybersecurity, or socio-technical probing.
β’ Capability to probe AI systems adversarially and push them to their limits.
β’ Familiarity with frameworks or benchmarks for systematic testing.
β’ Ability to communicate risks effectively to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects and clients.
β’ Preferred qualifications include specialties in adversarial machine learning, cybersecurity, socio-technical risks, or creative probing.
β’ Must operate as an independent contractor.
β’ H1-B and STEM OPT candidates will not be supported.
β’ Fully remote position that allows you to work on your own schedule.
β’ Receive weekly payments through Stripe or Wise based on services provided.
β’ Projects may be extended, shortened, or concluded early based on requirements and performance.
β’ Access to wellness resources and clear guidelines for projects involving higher sensitivity.
β’ Opportunity to gain experience in human data-driven AI red teaming at the forefront of safety.
β’ Play a direct role in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Competitive compensation.
β’ Referral bonuses of up to $100 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.