
AI Safety Experts, English β Bengali
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in United States.
β’ Conduct red-team evaluations of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.
β’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Implement taxonomies, benchmarks, and playbooks to ensure consistency in testing.
β’ Generate reproducible reports, datasets, and attack scenarios for clients.
β’ Analyze AI outputs related to sensitive issues such as bias, misinformation, and harmful behaviors.
β’ Discover vulnerabilities that automated assessments may overlook.
β’ Increase evaluation coverage while minimizing unexpected outcomes in production.
β’ Enhance customer AI systems through rigorous adversarial testing.
β’ Fluent or native proficiency in both English and Bengali.
β’ Previous experience in red teaming focused on AI adversarial tasks, cybersecurity, or socio-technical probing.
β’ Capability to challenge systems to their limits through adversarial testing.
β’ Proficiency in utilizing frameworks or benchmarks rather than employing random hacks.
β’ Ability to clearly communicate risks to both technical and non-technical stakeholders.
β’ Flexibility to adapt across various projects and client needs.
β’ Experience in adversarial machine learning is a desirable asset.
β’ Background in cybersecurity, such as penetration testing, exploit development, or reverse engineering, is a plus.
β’ Experience in managing socio-technical risks is considered an advantage.
β’ Background in creative probing through psychology, acting, or writing is a nice-to-have.
β’ Must be an independent contractor.
β’ H1-B and STEM OPT candidates are not eligible for this position.
β’ Fully remote position.
β’ Flexible scheduling options.
β’ Weekly payments processed via Stripe or Wise.
β’ Opportunity to gain experience in human data-driven AI red teaming.
β’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Collaboration with leading researchers in the field.
β’ Reasonable accommodations available upon request.
β’ Participation in higher-sensitivity projects is optional.
β’ Clear guidelines and wellness resources provided for handling sensitive content.
β’ Referral payments of up to $90 for each successful referral.
CVS Health
One Impression
Volga Partners
Mercor
Get handpicked remote jobs straight to your inbox weekly.