AI Safety Expert – English, Bengali

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Generate human data by annotating failures, categorizing vulnerabilities, and highlighting systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistency in testing.

β€’ Create reproducible reports, datasets, and attack cases for clients.

β€’ Review AI outputs that involve sensitive subjects, including bias, misinformation, and harmful behaviors.

β€’ Identify vulnerabilities that automated tests may overlook.

β€’ Broaden evaluation coverage and minimize unexpected issues in production.

β€’ Assist clients in enhancing the safety, robustness, and trustworthiness of AI systems.

β€’ Engage in projects focused on training and improving cutting-edge AI systems.


⛳️ Requirements

β€’ Fluent or native proficiency in English and Bengali.

β€’ Strong discernment regarding language and content.

β€’ Capability to evaluate whether AI responses are accurate, complete, and suitable, and articulate the reasoning behind assessments.

β€’ Meticulous attention to detail regarding subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Ability to clearly explain reasoning to both technical and non-technical audiences.

β€’ Flexibility across different projects, task types, and client needs.

β€’ Engagement as an independent contractor.

β€’ Not eligible as an H1-B or STEM OPT candidate.

β€’ Preferred expertise in areas such as adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.

β€’ Background or knowledge in jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction, penetration testing, exploit development, reverse engineering, harassment/disinformation probing, abuse analysis, conversational AI evaluation, psychology, acting, or innovative adversarial writing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule to accommodate your own timing.

β€’ Weekly payments via Stripe or Wise.

β€’ Optional participation in higher-sensitivity projects.

β€’ Clear guidelines and wellness resources for work involving sensitive content.

β€’ Competitive compensation.

β€’ Opportunity to collaborate with leading researchers.

β€’ Reasonable accommodations available upon request.

β€’ Referral bonuses of up to $90 for each successful referral, subject to limits.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board17 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers