AI Safety Expert – English, Swedish

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$48 – $62/hour

Posted Aug 24

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents, utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack scenarios for client use.

β€’ Evaluate AI outputs related to sensitive issues such as bias, misinformation, and harmful behaviors.

β€’ Broaden evaluation coverage and reveal vulnerabilities that automated tests may overlook.

β€’ Engage in projects aimed at training and enhancing cutting-edge AI systems.


⛳️ Requirements

β€’ Native or fluent proficiency in both English and Swedish.

β€’ Previous experience in red teaming within the realms of AI adversarial work, cybersecurity, or socio-technical probing.

β€’ Capability to probe systems adversarially and challenge them to their limits.

β€’ Experience with frameworks or benchmarks for systematic testing methodologies.

β€’ Proficiency in articulating risks clearly to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects and clientele.

β€’ Experience in adversarial machine learning, including but not limited to jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction (preferred).

β€’ Background in cybersecurity specialties such as penetration testing, exploit development, or reverse engineering (preferred).

β€’ Familiarity with socio-technical risk assessment, including harassment/disinformation probing, abuse analysis, or conversational AI evaluation (preferred).

β€’ Creative probing experience in fields like psychology, acting, or writing (preferred).

β€’ Candidates must not be on an H-1B or STEM OPT visa.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule allowing you to set your own hours.

β€’ Receive weekly payments via Stripe or Wise.

β€’ Gain valuable experience in human data-driven AI red teaming.

β€’ Play a direct role in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Competitive compensation.

β€’ Collaborate with leading researchers in the field.

β€’ Reasonable accommodations available upon request.

β€’ Earn referral bonuses of up to $250 for each successful referral (conditions may apply).

People also viewed

Mercor3 hours ago

Bilingual Dutch STEM Expert – AI Safety

NL flagNetherlands OnlyPart-timeArtificial Intelligence$61 – $65/hour
ApplyView job
Mercor4 hours ago

Bilingual Danish STEM Expert – AI Safety

DK flagDenmark OnlyPart-timeArtificial Intelligence$61 – $65/hour
ApplyView job
Mercor4 hours ago

Bilingual German Generalist Expert – AI Safety

DE flagGermany OnlyPart-timeArtificial Intelligence$48 – $52/hour
ApplyView job
Quality Digital5 hours ago

Junior Data and AI Analyst

BR flagBrazil OnlyFull-timeArtificial Intelligence
ApplyView job
Logic20/20, Inc.5 hours ago

Consulting Manager – AI Enablement

US flagWashington OnlyFull-timeArtificial Intelligence$157.2k – $177.9k/year
ApplyView job
Final Strike Games7 hours ago

AI Designer

US flagUnited States OnlyFull-timeArtificial Intelligence$90k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers