AI Safety Expert, English, Finnish

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$48 – $62/hour

Posted 10 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Conduct red team assessments on conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.

• Document failures, categorize vulnerabilities, and highlight systemic risks.

• Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures.

• Generate reproducible reports, datasets, and attack scenarios.

• Investigate sensitive AI outputs related to bias, misinformation, and harmful behaviors.

• Broaden evaluation scope and uncover vulnerabilities that automated tests may overlook.

• Work in collaboration with leading researchers to aid in the training and enhancement of advanced AI systems.


⛳️ Requirements

• Proficient in both English and Finnish.

• Previous experience in red teaming focused on AI adversarial efforts, cybersecurity, or socio-technical probing.

• Capability to adversarially probe AI systems and challenge them to their limits.

• Familiarity with frameworks or benchmarks for systematic testing.

• Competence in clearly communicating risks to both technical and non-technical audiences.

• Flexibility to adapt to various projects and clientele.

• Must have independent contractor status.

• Must not require H1-B or STEM OPT sponsorship.

• Desirable areas of expertise include adversarial ML, cybersecurity, socio-technical risk, or innovative probing techniques.


🏝️ Benefits

• Fully remote working arrangement.

• Flexible schedule tailored to your needs.

• Weekly payments processed through Stripe or Wise.

• Competitive compensation.

• Reasonable accommodations available upon request.

• Optional involvement in high-sensitivity projects.

• Clear guidelines and wellness resources provided for high-sensitivity projects.

• Referral bonuses of up to $250 for each successful referral.

People also viewed

Humana9 hours ago

Lead Solutions, AI Enablement Architect

US flagDistrict of Columbia, +8 more statesFull-timeArtificial Intelligence$142.3k – $195.7k/year
ApplyView job
Aviva9 hours ago

Conversational Designer – Artificial Intelligence

BR flagBrazil OnlyFull-timeArtificial Intelligence
ApplyView job
Madiff10 hours ago

AI Reverse Mentor

PL flagPoland OnlyFull-timeArtificial Intelligence
ApplyView job
Mercor10 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
Laivly11 hours ago

AI Deployment Manager

CA flagCanada OnlyFull-timeArtificial Intelligence
ApplyView job
Futurice12 hours ago

AI Director

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence£90k – £130k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers