
AI Safety Experts, English β Tamil
Posted 23 hours ago

Posted 23 hours ago
This is a fully remote position, open to applicants in United States.
β’ Engage in red-teaming for conversational AI models and agents through the use of jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
β’ Document failures, categorize vulnerabilities, and highlight systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures.
β’ Generate reproducible reports, datasets, and attack scenarios.
β’ Assess AI outputs related to sensitive subjects like bias, misinformation, and harmful behaviors.
β’ Identify vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues during production.
β’ Assist clients in enhancing the safety and robustness of their AI systems.
β’ Proficient/native fluency in both English and Tamil.
β’ Strong discernment regarding language and content, including assessing the accuracy, completeness, and appropriateness of AI responses.
β’ Capability to detect subtle errors, inconsistencies, and omissions.
β’ Consistent adherence to taxonomies, benchmarks, playbooks, guidelines, and quality standards.
β’ Ability to communicate reasoning clearly to both technical and non-technical audiences.
β’ Flexibility across various projects, task types, and client needs.
β’ Engagement as an independent contractor.
β’ Must not require H1-B or STEM OPT sponsorship.
β’ Preferred experience or specialization in adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.
β’ Fully remote position.
β’ Flexible work schedule, allowing you to manage your own time.
β’ Weekly payments processed via Stripe or Wise.
β’ Project durations may be adjusted based on needs and performance.
β’ Option to participate in higher-sensitivity projects is available.
β’ Access to clear guidelines and wellness resources for handling sensitive content.
β’ Reasonable accommodations provided upon request.
β’ Competitive compensation.
β’ Opportunity to collaborate with leading researchers in the field.
β’ Referral bonuses of up to $90 for each successful referral, subject to certain limits.
Epoch AI
Wagmo
Mercor
Blend360
Get handpicked remote jobs straight to your inbox weekly.