AI Safety Experts – English, Finnish

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$48 – $62/hour

Posted 14 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Conduct red-team assessments on conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

• Create human data by annotating failures, classifying vulnerabilities, and identifying systemic risks.

• Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures.

• Generate reproducible reports, datasets, and attack case studies for clients.

• Evaluate AI outputs related to sensitive issues, including bias, misinformation, and harmful behaviors.

• Identify vulnerabilities that automated tests may overlook.

• Increase evaluation coverage while minimizing unexpected issues in production.

• Engage in projects aimed at training and enhancing cutting-edge AI systems.


⛳️ Requirements

• Fluent or native proficiency in both English and Finnish.

• Previous experience in red teaming focused on AI adversarial work, cybersecurity, or socio-technical probing.

• Capability to test AI systems adversarially and push them to their limits.

• Experience with frameworks or benchmarks for structured testing methodologies.

• Proficiency in clearly communicating risks to both technical and non-technical audiences.

• Flexibility to adapt to various projects and client needs.

• Preferred expertise includes adversarial machine learning, cybersecurity, socio-technical risk, and innovative probing techniques.

• Position for independent contractors only.

• Must not require H1-B or STEM OPT sponsorship.


🏝️ Benefits

• Fully remote position.

• Flexible work schedule, allowing you to manage your own time.

• Weekly payments processed through Stripe or Wise.

• Project timelines may vary, with opportunities for extensions or early conclusion based on requirements and performance.

• Access to wellness resources and clear guidelines for projects with higher sensitivity.

• Reasonable accommodations available upon request.

• Chance to gain experience in human data-driven AI red teaming.

• Collaborate with leading experts in the field.

• Referral bonuses of up to $250 for each successful referral (certain limitations apply).

People also viewed

Deepgram4 hours ago

Research Staff – Voice AI Foundations

AU flagAustralia OnlyFull-timeArtificial Intelligence
ApplyView job
Deepgram4 hours ago

Research Staff, Voice AI Foundations

NL flagNetherlands OnlyFull-timeArtificial Intelligence
ApplyView job
Deepgram4 hours ago

Research Staff, Voice AI Foundations

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job
Mercor5 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
Meiks Affiliate Tipps17 hours ago

Agency Director – Scale with AI, Systems & Clear Communication

DE flagGermany OnlyFull-timeArtificial Intelligence€7,500 – €18.5k/month
ApplyView job
Keyrus17 hours ago

AI Business Value Advisor

CO flagColombia, +1 more countryFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers