AI Safety Expert, English, Dutch

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$48 – $62/hour

Posted Sep 3

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red teaming on conversational AI models and agents by exploring jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human data through annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Analyze AI outputs related to sensitive issues like bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Broaden evaluation coverage and minimize unexpected outcomes in production.

β€’ Assist Mercor clients in enhancing the safety, robustness, and trustworthiness of their AI systems.


⛳️ Requirements

β€’ Native fluency in both English and Dutch is essential.

β€’ Previous experience in red teaming within AI adversarial contexts, cybersecurity, or socio-technical probing is required.

β€’ Capability to probe systems adversarially and push them to their limits.

β€’ Experience with frameworks or benchmarks for structured testing is necessary.

β€’ Proficient in articulating risks to both technical and non-technical stakeholders.

β€’ Flexibility to adapt to various projects and clients.

β€’ Nice-to-have: Experience in adversarial machine learning, including jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction.

β€’ Nice-to-have: Experience in cybersecurity, such as penetration testing, exploit development, or reverse engineering.

β€’ Nice-to-have: Experience with socio-technical risks, including harassment/disinformation probing, abuse analysis, or testing conversational AI.

β€’ Nice-to-have: Creative probing experience in fields like psychology, acting, or writing.

β€’ Must be engaged as an independent contractor.

β€’ H1-B and STEM OPT candidates are not supported.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule.

β€’ Weekly payments via Stripe or Wise based on delivered services.

β€’ Project timelines can be adjusted, whether extended, shortened, or concluded early based on requirements and performance.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources available for handling sensitive content.

β€’ Competitive compensation.

β€’ Referral bonus of up to $250 for each successful referral.

People also viewed

WON.ai19 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board20 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor21 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor21 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor21 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner21 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers