AI Safety Expert – English, Marathi

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Engage in red-team activities for conversational AI models and agents, focusing on jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Document failures, categorize vulnerabilities, and identify systemic risks.

β€’ Adhere to taxonomies, benchmarks, and playbooks to ensure consistency in testing processes.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Review AI outputs related to sensitive issues such as bias, misinformation, and harmful behaviors.

β€’ Increase evaluation breadth and reveal vulnerabilities that automated testing might overlook.

β€’ Collaborate on initiatives aimed at training and improving frontier AI systems for Mercor's clientele.


⛳️ Requirements

β€’ Required fluency/native proficiency in English and Marathi.

β€’ Strong discernment regarding language and content; ability to evaluate the accuracy, completeness, and appropriateness of AI responses, along with the capacity to articulate reasoning.

β€’ Meticulous attention to subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Capability to communicate reasoning effectively to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Independent contractor status is required.

β€’ Candidates must not hold H-1B or STEM OPT status.

β€’ Preferred qualifications include expertise in adversarial ML, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Work remotely from anywhere.

β€’ Enjoy a flexible schedule; manage your own work hours.

β€’ Receive weekly payments through Stripe or Wise.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Access clear guidelines and wellness resources for higher-sensitivity initiatives.

β€’ Competitive compensation.

β€’ Opportunity to earn up to $90 for each successful referral (with referral limits applicable).

β€’ Reasonable accommodations available upon request.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers