AI Safety Expert – English, Malay

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$17 – $25/hour

Posted Sep 15

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments of conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human data by annotating deficiencies, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistency in testing.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Evaluate AI outputs related to sensitive subjects, including bias, misinformation, and harmful behaviors.

β€’ Assist in broadening evaluation coverage and identifying vulnerabilities that automated tests may overlook.

β€’ Collaborate with top researchers and play a role in training and improving advanced AI systems.


⛳️ Requirements

β€’ Required fluent/native proficiency in both English and Malay.

β€’ Previous experience in red teaming, particularly in AI adversarial work, cybersecurity, or socio-technical probing.

β€’ Capability to probe AI systems adversarially and push them to their limits.

β€’ Familiarity with frameworks or benchmarks for systematic testing.

β€’ Ability to communicate risks effectively to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects and clients.

β€’ Preferred qualifications include specialties in adversarial machine learning, cybersecurity, socio-technical risks, or creative probing.

β€’ Must operate as an independent contractor.

β€’ H1-B and STEM OPT candidates will not be supported.


🏝️ Benefits

β€’ Fully remote position that allows you to work on your own schedule.

β€’ Receive weekly payments through Stripe or Wise based on services provided.

β€’ Projects may be extended, shortened, or concluded early based on requirements and performance.

β€’ Access to wellness resources and clear guidelines for projects involving higher sensitivity.

β€’ Opportunity to gain experience in human data-driven AI red teaming at the forefront of safety.

β€’ Play a direct role in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Competitive compensation.

β€’ Referral bonuses of up to $100 for each successful referral.

People also viewed

WON.ai17 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board19 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner19 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers