AI Safety Expert – English, Bengali

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments of conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Create human data by annotating failures, identifying vulnerabilities, and highlighting systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Deliver reproducible reports, datasets, and case studies for clients.

β€’ Evaluate AI outputs related to sensitive issues, including bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Enhance evaluation coverage to minimize production surprises.

β€’ Fortify customer AI systems, contributing to the development of safer, more resilient, and trustworthy AI technologies.


⛳️ Requirements

β€’ Proficient/native fluency in English and Bengali is mandatory.

β€’ Strong judgment regarding language and content; capability to evaluate whether AI responses are accurate, comprehensive, and suitable.

β€’ Skill in recognizing subtle errors, inconsistencies, and gaps.

β€’ Ability to consistently adhere to guidelines and quality standards.

β€’ Competence in clearly articulating reasoning to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Engagement as an independent contractor is required.

β€’ Candidates should not be on an H1-B visa or STEM OPT program.

β€’ Preferred qualifications include specialties in adversarial ML, jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction, penetration testing, exploit development, reverse engineering, harassment/disinformation probing, abuse analysis, conversational AI testing, psychology, acting, or innovative adversarial writing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule; complete tasks at your convenience.

β€’ Weekly payments through Stripe or Wise based on services rendered.

β€’ Project duration may be extended, shortened, or concluded early based on needs and performance.

β€’ Reasonable accommodations available upon request.

β€’ Chance to gain experience in human data-driven AI red teaming.

β€’ Collaborate with top researchers in the field.

β€’ Referral bonuses of up to $90 for each successful referral, subject to limits.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers