
AI Safety Expert β English, Tamil
Posted 11 hours ago

Posted 11 hours ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team assessments of conversational AI models and agents using jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.
β’ Create human data by annotating errors, categorizing vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Evaluate AI outputs related to sensitive subjects such as bias, misinformation, or harmful behaviors.
β’ Identify vulnerabilities that automated testing may overlook.
β’ Provide reproducible artifacts that enhance the AI systems of customers.
β’ Broaden evaluation coverage and minimize unexpected issues in production.
β’ Proficiency in English and Tamil is essential.
β’ Native fluency in both English and Tamil is a must.
β’ Strong judgment regarding language and content is necessary.
β’ Capability to evaluate the accuracy, completeness, and appropriateness of AI responses and articulate the reasoning behind assessments.
β’ Ability to detect subtle errors, inconsistencies, and gaps in content.
β’ Consistent adherence to guidelines and quality standards is required.
β’ Skill in clearly explaining reasoning to both technical and non-technical audiences.
β’ Flexibility to adapt to various projects, task types, and client needs.
β’ Independent contractor position.
β’ Support for H1-B and STEM OPT candidates is not available.
β’ Preferred areas of expertise include adversarial machine learning, cybersecurity, socio-technical risk, and creative probing.
β’ This role is fully remote and can be performed according to your own schedule.
β’ Weekly payments issued via Stripe or Wise based on the services provided.
β’ Optional participation in higher-sensitivity projects.
β’ Access to clear guidelines and wellness resources for projects involving sensitive content.
β’ Competitive compensation.
β’ Opportunity to collaborate with leading researchers.
β’ Referral program that offers up to $90 for each successful referral.
Mercor
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.