
AI Safety Expert β English, Telugu
Posted 4 hours ago

Posted 4 hours ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team evaluations of conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.
β’ Document failures, categorize vulnerabilities, and identify systemic risks.
β’ Adhere to established taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Assess AI outputs related to sensitive subjects, including bias, misinformation, or harmful behaviors.
β’ Identify vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unforeseen issues during production.
β’ Enhance customer AI systems through consistent adversarial testing.
β’ Native proficiency in English and Telugu is essential.
β’ Strong discernment regarding language and content, including evaluating the accuracy, completeness, and appropriateness of AI responses.
β’ Ability to articulate reasoning clearly to both technical and non-technical audiences.
β’ Meticulous attention to detail, spotting subtle errors, inconsistencies, and omissions.
β’ Capacity to consistently adhere to guidelines and maintain quality standards.
β’ Flexibility to adapt across various projects, task types, and client needs.
β’ Engagement as an independent contractor is required.
β’ Must be capable of working without access to confidential or proprietary information from any employer, client, or institution.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Enjoy a fully remote position that allows you to work on your own schedule.
β’ Receive weekly payments through Stripe or Wise based on services provided.
β’ Option to participate in higher-sensitivity projects is available.
β’ Access to clear guidelines and wellness resources for handling higher-sensitivity content.
β’ Gain experience in human data-driven AI red teaming.
β’ Play a crucial role in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Collaborate with leading researchers in the field.
β’ Earn referral payments of up to $90 for each successful referral, subject to certain limits.
Mercor
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.