AI Safety Expert – English, Odia

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team evaluations of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

β€’ Create human data by annotating errors, categorizing vulnerabilities, and identifying systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack cases for customers to utilize.

β€’ Assess AI outputs related to sensitive subjects such as bias, misinformation, and harmful behaviors.

β€’ Identify vulnerabilities that automated testing fails to detect.

β€’ Increase evaluation coverage and minimize unexpected outcomes in production.

β€’ Fortify customer AI systems and promote the development of safer, more reliable AI technologies.

β€’ Collaborate with top researchers on initiatives aimed at training and improving cutting-edge AI systems.


⛳️ Requirements

β€’ Native proficiency in both English and Odia is essential.

β€’ Strong judgment regarding language and content; ability to evaluate whether AI responses are accurate, complete, and suitable, and articulate the reasoning behind assessments.

β€’ Keen attention to detail to identify subtle mistakes, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Capability to communicate reasoning effectively to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and clientele.

β€’ Candidates must not require H1-B or STEM OPT sponsorship.

β€’ Status as an independent contractor is required.

β€’ Preferred: experience in adversarial machine learning, encompassing jailbreak datasets, prompt injection, RLHF/DPO attacks, and model extraction.

β€’ Preferred: background in cybersecurity, including penetration testing, exploit development, and reverse engineering.

β€’ Preferred: experience in socio-technical risks, such as harassment/disinformation probing, abuse analysis, and conversational AI testing.

β€’ Preferred: background in psychology, acting, or writing for innovative adversarial thinking.


🏝️ Benefits

β€’ Fully remote position that allows you to work on your own schedule.

β€’ Optional involvement in higher-sensitivity projects.

β€’ Access to clear guidelines and wellness resources for projects involving sensitive content.

β€’ Weekly payments processed via Stripe or Wise based on services rendered.

β€’ Competitive compensation.

β€’ Reasonable accommodations available upon request.

β€’ Referral bonuses of up to $90 for each successful referral, subject to certain limits.

β€’ Opportunity to work alongside leading researchers.

β€’ Chance to gain experience in human data-driven AI red teaming.

β€’ Engagement as an independent contractor.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board17 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers