
AI Safety Expert β English, Telugu
Posted 13 hours ago

Posted 13 hours ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments on conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.
β’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing methodologies.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Evaluate AI outputs related to sensitive subjects, including bias, misinformation, or harmful behaviors.
β’ Identify vulnerabilities that automated testing may overlook.
β’ Enhance evaluation coverage and minimize unexpected issues during production.
β’ Fortify customer AI systems through adversarial testing.
β’ Native fluency in both English and Telugu is essential.
β’ Strong judgment regarding language and content quality.
β’ Ability to evaluate whether AI responses are accurate, complete, and appropriate, along with the capacity to articulate reasoning.
β’ Meticulous attention to detail, including subtle errors, inconsistencies, and gaps.
β’ Consistent adherence to guidelines and quality standards.
β’ Effective communication of reasoning to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, task types, and client needs.
β’ Status as an independent contractor is required.
β’ Must not require H1-B or STEM OPT sponsorship.
β’ Preferred areas of expertise include adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.
β’ Fully remote position.
β’ Flexible work schedule; manage your own time.
β’ Weekly payments via Stripe or Wise.
β’ Project durations may vary based on needs and performance.
β’ Access to wellness resources and clear guidelines for projects with higher sensitivity.
β’ Opportunity to gain experience in human data-driven AI red teaming.
β’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Competitive compensation.
β’ Collaboration opportunities with leading researchers.
β’ Reasonable accommodations available upon request.
Mercor
snipKI
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.