AI Safety Expert – English, Kannada

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 2 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

β€’ Evaluate AI outputs related to sensitive subjects such as bias, misinformation, or harmful behaviors.

β€’ Document failures, categorize vulnerabilities, and identify systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Broaden evaluation coverage to minimize unexpected issues during production.

β€’ Assist Mercor clients in enhancing the robustness, safety, and trustworthiness of their AI systems.


⛳️ Requirements

β€’ Proficiency in both English and Kannada is mandatory.

β€’ Strong judgment regarding language and content issues.

β€’ Capacity to evaluate the accuracy, completeness, and appropriateness of AI responses.

β€’ Ability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Meticulous attention to detail, particularly concerning subtle errors, inconsistencies, and omissions.

β€’ Capability to adhere to guidelines, taxonomies, benchmarks, playbooks, and quality standards without fail.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Status as an independent contractor is required.

β€’ H1-B and STEM OPT applicants will not be considered.

β€’ Preferred expertise includes adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Position is fully remote.

β€’ Offers a flexible schedule allowing you to work at your convenience.

β€’ Receive weekly payments through Stripe or Wise.

β€’ Competitive compensation package.

β€’ Access to wellness resources and clear protocols for high-sensitivity projects.

β€’ Reasonable accommodations available upon request.

β€’ Referral program: earn up to $90 for each successful referral.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers