
AI Safety Expert, English, Telugu
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments on conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Document failures, categorize vulnerabilities, and highlight systemic risks.
β’ Implement taxonomies, benchmarks, and playbooks to ensure consistent testing.
β’ Generate reproducible reports, datasets, and attack cases for our customers.
β’ Evaluate AI outputs concerning sensitive subjects like bias, misinformation, or harmful behaviors.
β’ Assist Mercor clients in enhancing the robustness, safety, and trustworthiness of their AI systems.
β’ Fluency in both English and Telugu is mandatory.
β’ Strong judgment regarding language and content; capability to evaluate whether AI responses are accurate, comprehensive, and appropriate.
β’ Proficient in identifying subtle errors, inconsistencies, and gaps.
β’ Consistent adherence to taxonomies, benchmarks, playbooks, guidelines, and quality standards.
β’ Ability to clearly articulate reasoning to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, task types, and customer needs.
β’ Status as an independent contractor is required.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Fully remote position with the flexibility to set your own schedule.
β’ Weekly payments via Stripe or Wise.
β’ Access to wellness resources and clear guidelines for projects with higher sensitivity.
β’ Competitive compensation.
β’ Reasonable accommodations available upon request.
β’ Referral bonuses of up to $90 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.