
AI Safety Expert, English β Bengali
Posted 11 hours ago

Posted 11 hours ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team evaluations of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.
β’ Create human data by annotating failures, identifying vulnerabilities, and highlighting systemic risks.
β’ Adhere to established taxonomies, benchmarks, and playbooks to maintain consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Examine AI outputs related to sensitive subjects, including bias, misinformation, and harmful behaviors.
β’ Identify vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues in production.
β’ Assist clients in enhancing the safety, robustness, and trustworthiness of their AI systems.
β’ Native proficiency in both English and Bengali.
β’ Strong discernment regarding language and content.
β’ Capability to evaluate the accuracy, completeness, and appropriateness of AI responses, along with the ability to articulate reasoning.
β’ Ability to detect subtle errors, inconsistencies, and omissions.
β’ Consistent adherence to guidelines and quality standards.
β’ Skill in clearly communicating reasoning to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, task types, and client needs.
β’ Engagement as an independent contractor.
β’ Must not be an H-1B or STEM OPT candidate.
β’ Preferred areas of expertise: adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.
β’ Fully remote position.
β’ Flexible working hours, allowing you to set your own schedule.
β’ Weekly compensation through Stripe or Wise.
β’ Projects may be adjusted in duration or concluded early based on requirements and performance.
β’ Participation in higher-sensitivity projects is optional.
β’ Access to clear guidelines and wellness resources for higher-sensitivity projects.
β’ Competitive compensation.
β’ Opportunity to collaborate with leading researchers.
β’ Reasonable accommodations available upon request.
β’ Referral program with earnings of up to $90 for each successful referral (referral limits apply).
Gartner
Mercor
manara
Escalent
Get handpicked remote jobs straight to your inbox weekly.