AI Safety Expert – English, Odia

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 3 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents by utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Produce human data through the annotation of failures, classification of vulnerabilities, and identification of systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Evaluate AI outputs related to sensitive topics like bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated testing methods may overlook.

β€’ Enhance evaluation coverage and minimize unexpected issues during production.

β€’ Assist clients in enhancing the safety and robustness of their AI systems.


⛳️ Requirements

β€’ Proficient fluency in both English and Odia is essential.

β€’ Strong judgment regarding language and content; capability to evaluate whether AI responses are accurate, comprehensive, and suitable, along with the ability to articulate the rationale.

β€’ Keen ability to detect subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Capability to clearly communicate reasoning to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Engagement as an independent contractor.

β€’ Must not need H1-B or STEM OPT sponsorship.

β€’ Preferred: experience in adversarial machine learning, including jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction.

β€’ Preferred: experience in cybersecurity, including penetration testing, exploit development, or reverse engineering.

β€’ Preferred: experience in socio-technical risk, including harassment/disinformation probing, abuse analysis, or conversational AI testing.

β€’ Preferred: background in psychology, acting, or writing to foster unconventional adversarial thinking.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible working hours.

β€’ Weekly payments via Stripe or Wise.

β€’ Project durations may vary based on needs and performance.

β€’ Access to wellness resources and clear guidelines for sensitive projects.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Collaboration with leading researchers in the field.

β€’ Competitive compensation.

β€’ Referral bonuses of up to $90 for each successful referral (referral limits apply).

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board17 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers