AI Safety Expert – English, Tamil

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Evaluate AI outputs related to sensitive subjects such as bias, misinformation, and harmful behaviors.

β€’ Document failures, categorize vulnerabilities, and identify systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing processes.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Investigate AI systems to reveal vulnerabilities that automated tests might overlook.

β€’ Broaden evaluation coverage to minimize unexpected outcomes in production.

β€’ Assist Mercor clients in enhancing the safety, robustness, and reliability of their AI systems.


⛳️ Requirements

β€’ Proficiency in English and Tamil is mandatory.

β€’ Strong discernment regarding language and content, including assessing whether AI responses are accurate, complete, and suitable.

β€’ Ability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Meticulous attention to detail, particularly with respect to subtle errors, inconsistencies, and omissions.

β€’ Consistent adherence to taxonomies, benchmarks, playbooks, guidelines, and quality standards.

β€’ Flexibility to adapt across various projects, task types, and clientele.

β€’ Engagement as an independent contractor is required.

β€’ Must be able to work without access to confidential or proprietary information from any other employer, client, or institution.

β€’ H1-B and STEM OPT candidates are not eligible.

β€’ Preferred experience in adversarial machine learning, cybersecurity, socio-technical risk, or creative probing is a plus.


🏝️ Benefits

β€’ Fully remote position with flexible, self-directed scheduling.

β€’ Weekly payments through Stripe or Wise based on services performed.

β€’ Project durations may be extended, shortened, or concluded early based on needs and performance.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources are available for work involving sensitive content.

β€’ Competitive compensation.

β€’ Opportunity to collaborate with top-tier researchers.

β€’ Referral bonuses of up to $90 for each successful referral, subject to certain limits.

β€’ Reasonable accommodations available upon request.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers