AI Safety Expert – English, Telugu

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 5 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

β€’ Document failures, categorize vulnerabilities, and identify systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing processes.

β€’ Generate reproducible reports, datasets, and attack case studies for clients.

β€’ Analyze AI outputs related to sensitive subjects such as bias, misinformation, and harmful behaviors.

β€’ Collaborate on initiatives aimed at training and enhancing AI systems.

β€’ Identify vulnerabilities that automated testing might overlook.

β€’ Broaden evaluation coverage and minimize unexpected issues in production.


⛳️ Requirements

β€’ Required fluent/native proficiency in both English and Telugu.

β€’ Strong discernment regarding language and content, including assessing the accuracy, completeness, and appropriateness of AI responses.

β€’ Ability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Meticulous attention to subtle errors, inconsistencies, and omissions.

β€’ Consistently adhere to guidelines and quality standards.

β€’ Flexibility across various projects, task types, and clientele.

β€’ Engagement as an independent contractor.

β€’ Candidates must not hold H1-B or STEM OPT status.

β€’ Preferred specialties include adversarial machine learning, jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction, penetration testing, exploit development, reverse engineering, harassment/disinformation investigation, abuse analysis, conversational AI testing, psychology, acting, and writing.


🏝️ Benefits

β€’ Fully remote position that allows you to work on your own schedule.

β€’ Weekly payments through Stripe or Wise based on services performed.

β€’ Opportunity to gain experience in human data-driven AI red teaming at the cutting edge of safety.

β€’ Play a direct role in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources available for higher-sensitivity projects.

β€’ Reasonable accommodations provided upon request.

β€’ Referral bonuses of up to $90 for each successful referral, with no specified limit on the number of referrals.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers