AI Safety Experts, English, Kannada

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted Sep 19

This is a fully remote position, open to applicants in United States.

📋 Description

• Conduct red-team exercises on conversational AI models and agents by utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

• Produce human data by annotating failures, classifying vulnerabilities, and identifying systemic risks.

• Implement taxonomies, benchmarks, and playbooks to ensure consistent testing methodologies.

• Generate reproducible reports, datasets, and attack case studies.

• Analyze AI outputs related to sensitive subjects like bias, misinformation, and harmful behaviors.

• Broaden evaluation coverage and discover vulnerabilities that automated tests may overlook.

• Develop artifacts that enhance the robustness of customers’ AI systems.


⛳️ Requirements

• Native proficiency in English and Kannada.

• Strong judgment regarding language and content, particularly in assessing the accuracy, completeness, and appropriateness of AI responses.

• Ability to clearly articulate reasoning to both technical and non-technical audiences.

• Meticulous attention to detail, identifying subtle errors, inconsistencies, and gaps.

• Capability to consistently adhere to guidelines and quality standards.

• Flexibility to adapt across various projects, task types, and client needs.

• Independent contractor status required.

• Must not hold H1-B or STEM OPT status.

• Preferred qualifications include expertise in adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

• Fully remote position.

• Flexible schedule, allowing work to be completed at your convenience.

• Weekly payments via Stripe or Wise.

• Project durations may be extended, shortened, or concluded early based on requirements and performance.

• Opportunity to gain experience in human data-driven AI red teaming.

• Play a direct role in enhancing the robustness, safety, and trustworthiness of AI systems.

• Collaborate with top researchers in the field.

• Referral bonuses of up to $90 for each successful referral, with no cap on the number of referrals (certain restrictions may apply).

• Reasonable accommodations available upon request.

People also viewed

WON.ai17 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner19 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers