AI Safety Expert, English – Tamil

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 6 days ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Conduct red-team assessments on conversational AI models and agents by employing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

• Create high-quality human data by annotating failures, classifying vulnerabilities, and identifying systemic risks.

• Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing processes.

• Develop reproducible reports, datasets, and attack scenarios for clients.

• Evaluate AI outputs concerning sensitive topics like bias, misinformation, and harmful behaviors.

• Identify vulnerabilities that automated tests might overlook.

• Broaden evaluation coverage to minimize surprises in production.

• Assist clients in enhancing the safety, robustness, and reliability of their AI systems.


⛳️ Requirements

• Proficiency in English and Tamil is essential.

• Strong discernment regarding language and content.

• Capability to assess whether AI responses are accurate, complete, and suitable.

• Ability to articulate reasoning clearly to both technical and non-technical audiences.

• Meticulous attention to subtle errors, inconsistencies, and omissions.

• Consistent adherence to guidelines and quality standards.

• Flexibility across various projects, task types, and clients.

• Independent contractor status is required.

• Candidates must not hold H1-B or STEM OPT status.

• Preferred qualifications: experience in adversarial ML, including familiarity with jailbreak datasets, prompt injection, RLHF/DPO attacks, and model extraction.

• Preferred qualifications: background in cybersecurity, including penetration testing, exploit development, and reverse engineering.

• Preferred qualifications: experience in socio-technical risk, including harassment/disinformation probing, abuse analysis, and conversational AI testing.

• Preferred qualifications: background in psychology, acting, or writing to foster unconventional adversarial thinking.


🏝️ Benefits

• Fully remote position.

• Flexible working hours; complete tasks on your preferred schedule.

• Weekly payments through Stripe or Wise.

• Competitive salary.

• Participation in higher-sensitivity projects is optional.

• Clear content guidelines and access to wellness resources.

• Reasonable accommodations available upon request.

• Referral bonuses of up to $90 for successful referrals.

People also viewed

WON.ai17 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner19 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers