AI Safety Expert, English, Portuguese

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$29 – $45/hour

Posted Aug 24

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Engage in red teaming for conversational AI models and agents

β€’ Conduct tests for jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation

β€’ Create human data by annotating failures, identifying vulnerabilities, and highlighting systemic risks

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures

β€’ Generate reproducible reports, datasets, and attack scenarios for our clients

β€’ Investigate sensitive issues including bias, misinformation, and harmful behaviors

β€’ Assist in broadening evaluation coverage and minimizing production surprises

β€’ Collaborate with top researchers on projects aimed at training and enhancing AI systems


⛳️ Requirements

β€’ Native proficiency in both English and Portuguese (global, excluding Brazilian Portuguese)

β€’ Previous experience in red teaming, particularly in AI adversarial work, cybersecurity, or socio-technical investigations

β€’ Capability to probe AI models through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation

β€’ Proficiency in annotating failures, classifying vulnerabilities, and flagging systemic risks

β€’ Familiarity with adhering to taxonomies, benchmarks, and playbooks

β€’ Ability to produce reproducible reports, datasets, and attack cases

β€’ Skill in articulating risks clearly to both technical and non-technical audiences

β€’ Flexibility to adapt across various projects and clients

β€’ H1-B and STEM OPT candidates are not accepted


🏝️ Benefits

β€’ Fully remote position

β€’ Flexible, self-directed work schedule

β€’ Weekly payments through Stripe or Wise

β€’ Participation in higher-sensitivity projects is optional

β€’ Access to clear content guidelines and wellness resources

β€’ Opportunity to gain experience in human data-driven AI red teaming

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems

β€’ Competitive salary

β€’ Collaboration with leading researchers in the field

β€’ Referral bonuses of up to $180 for each successful referral

People also viewed

Mercor18 hours ago

Bilingual Dutch STEM Expert – AI Safety

NL flagNetherlands OnlyPart-timeArtificial Intelligence$61 – $65/hour
ApplyView job
Mercor19 hours ago

Bilingual Danish STEM Expert – AI Safety

DK flagDenmark OnlyPart-timeArtificial Intelligence$61 – $65/hour
ApplyView job
Mercor19 hours ago

Bilingual German Generalist Expert – AI Safety

DE flagGermany OnlyPart-timeArtificial Intelligence$48 – $52/hour
ApplyView job
Quality Digital19 hours ago

Junior Data and AI Analyst

BR flagBrazil OnlyFull-timeArtificial Intelligence
ApplyView job
Logic20/20, Inc.19 hours ago

Consulting Manager – AI Enablement

US flagWashington OnlyFull-timeArtificial Intelligence$157.2k – $177.9k/year
ApplyView job
Final Strike Games21 hours ago

AI Designer

US flagUnited States OnlyFull-timeArtificial Intelligence$90k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers