AI Safety Expert – English, Odia

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents through jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

β€’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure testing consistency.

β€’ Generate reproducible reports, datasets, and attack case studies.

β€’ Evaluate AI outputs concerning sensitive issues such as bias, misinformation, and harmful behaviors.

β€’ Identify vulnerabilities that automated tests overlook.

β€’ Broaden evaluation coverage and minimize surprises in production.

β€’ Enhance customer AI systems and aid in creating safer, more reliable AI solutions.

β€’ Collaborate with prominent researchers on initiatives focused on training and improving cutting-edge AI models.


⛳️ Requirements

β€’ Proficient/native fluency in both English and Odia.

β€’ Strong discernment regarding language and content; capable of evaluating the accuracy, completeness, and appropriateness of AI responses, along with the ability to articulate the reasons behind assessments.

β€’ Keen observation skills for identifying subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Ability to articulate reasoning effectively to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, tasks, and client needs.

β€’ Engagement as an independent contractor.

β€’ Currently unable to support H1-B or STEM OPT candidates.

β€’ Preferred areas of expertise: adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Fully remote position that allows for flexible scheduling.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources available for managing sensitive content.

β€’ Weekly payments processed through Stripe or Wise based on services rendered.

β€’ Competitive compensation package.

β€’ Referral bonuses of up to $90 for each successful referral, subject to referral limits.

β€’ Reasonable accommodations available upon request.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers