AI Safety Expert – English, Telugu

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 5 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Document failures, categorize vulnerabilities, and identify systemic risks.

β€’ Adhere to established taxonomies, benchmarks, and playbooks to ensure consistency in testing.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Analyze AI outputs related to sensitive topics, including bias, misinformation, or harmful behaviors.

β€’ Broaden evaluation coverage and discover vulnerabilities that automated tests may overlook.

β€’ Collaborate on initiatives aimed at training and improving cutting-edge AI systems.


⛳️ Requirements

β€’ Native fluency in both English and Telugu is essential.

β€’ Strong judgment regarding language and content, especially in evaluating the accuracy, completeness, and appropriateness of AI responses.

β€’ Capacity to detect subtle errors, inconsistencies, and omissions.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Ability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Engagement as an independent contractor.

β€’ Must not require H1-B or STEM OPT sponsorship.

β€’ Preferred expertise includes adversarial machine learning, cybersecurity, socio-technical risk, and creative probing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible working hours, allowing you to manage your own schedule.

β€’ Weekly payments through Stripe or Wise.

β€’ Project durations may be extended, shortened, or concluded early based on requirements and performance.

β€’ Competitive compensation.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Reasonable accommodations available upon request.

β€’ Referral bonus of up to $90 for each successful referral.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers