AI Safety Expert, English, Portuguese

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$29 – $45/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red team assessments on conversational AI models and agents using techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Generate human data by annotating failures, identifying vulnerabilities, and flagging systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Create reproducible reports, datasets, and attack scenarios for clients.

β€’ Evaluate AI outputs related to sensitive issues such as bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Increase evaluation coverage and minimize unexpected issues in production.

β€’ Enhance customer AI systems through adversarial testing.


⛳️ Requirements

β€’ Native proficiency in English and Portuguese (global, excluding Brazilian Portuguese).

β€’ Previous experience with red teaming in AI adversarial contexts, cybersecurity, or socio-technical probing.

β€’ Capability to challenge AI systems adversarially and push them to their limits.

β€’ Proficiency in using frameworks or benchmarks for structured testing methodologies.

β€’ Skill in articulating risks clearly to both technical and non-technical audiences.

β€’ Flexibility to adapt to various projects and client needs.

β€’ Independent contractor status required.

β€’ Must not need H1-B or STEM OPT sponsorship.

β€’ Nice-to-have expertise: adversarial machine learning, cybersecurity, socio-technical risk analysis, or innovative probing techniques.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule allowing you to manage your own hours.

β€’ Weekly payments processed via Stripe or Wise.

β€’ Project timelines may be adjusted, shortened, or concluded early based on requirements and performance.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources available for projects involving sensitive content.

β€’ Competitive salary.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Collaborate with top researchers in the field.

β€’ Reasonable accommodations can be made upon request.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers