AI Safety Experts – English, Assamese

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 16 hours ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments of conversational AI models and agents through jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Analyze AI outputs related to sensitive topics, including bias, misinformation, and harmful behaviors.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Broaden evaluation coverage and minimize unexpected outcomes in production.

β€’ Engage in projects focused on training and enhancing advanced AI systems.


⛳️ Requirements

β€’ Native proficiency in English and Assamese is essential.

β€’ Strong discernment regarding language and content quality.

β€’ Capability to evaluate whether AI responses are accurate, complete, and suitable.

β€’ Proficiency in articulating reasoning clearly to both technical and non-technical audiences.

β€’ Meticulous attention to detail, including subtle errors, inconsistencies, and gaps.

β€’ Ability to consistently adhere to guidelines, taxonomies, benchmarks, playbooks, and quality standards.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Must be an independent contractor.

β€’ Currently, we are unable to accommodate H1-B or STEM OPT candidates.

β€’ Preferred qualifications include expertise in adversarial ML, cybersecurity, socio-technical risk, conversational AI testing, psychology, acting, and writing.


🏝️ Benefits

β€’ Fully remote working environment.

β€’ Flexible scheduling options.

β€’ Weekly payments through Stripe or Wise.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Collaboration with top researchers in the field.

β€’ Reasonable accommodations available upon request.

β€’ Referral bonuses of up to $90 for each successful referral, subject to limits.

People also viewed

WON.ai14 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board15 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor16 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor16 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner16 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job
Designlab17 hours ago

Instructor, AI Workflows – Agents

US flagNew York OnlyPart-timeArtificial Intelligence$90 – $120/hour
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers