AI Safety Expert – English, Telugu

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 13 hours ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing methodologies.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Evaluate AI outputs related to sensitive subjects, including bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated testing may overlook.

β€’ Enhance evaluation coverage and minimize unexpected issues during production.

β€’ Fortify customer AI systems through adversarial testing.


⛳️ Requirements

β€’ Native fluency in both English and Telugu is essential.

β€’ Strong judgment regarding language and content quality.

β€’ Ability to evaluate whether AI responses are accurate, complete, and appropriate, along with the capacity to articulate reasoning.

β€’ Meticulous attention to detail, including subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Effective communication of reasoning to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Status as an independent contractor is required.

β€’ Must not require H1-B or STEM OPT sponsorship.

β€’ Preferred areas of expertise include adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule; manage your own time.

β€’ Weekly payments via Stripe or Wise.

β€’ Project durations may vary based on needs and performance.

β€’ Access to wellness resources and clear guidelines for projects with higher sensitivity.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Competitive compensation.

β€’ Collaboration opportunities with leading researchers.

β€’ Reasonable accommodations available upon request.

People also viewed

Mercor13 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
snipKI20 hours ago

Community Manager – AI Community

DE flagGermany OnlyPart-timeArtificial Intelligence
ApplyView job
Mercor22 hours ago

AI Safety Expert – English, Finnish

US flagUnited States OnlyFreelanceArtificial Intelligence$48 – $62/hour
ApplyView job
Mercor22 hours ago

AI Safety Expert – English, Tamil

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor22 hours ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job
Mercor1 day ago

AI Safety Red Teamer

US flagUnited States OnlyFreelanceArtificial Intelligence$70 – $84/hour
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers