
AI Safety Expert β English, Assamese
Posted Aug 21

Posted Aug 21
This is a fully remote position, open to applicants in United States.
β’ Engage in red-team evaluations of conversational AI models and agents.
β’ Conduct tests for jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.
β’ Create high-quality human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to maintain consistency in testing.
β’ Prepare reproducible reports, datasets, and attack cases for clients.
β’ Investigate AI systems for vulnerabilities that automated tests may overlook.
β’ Broaden evaluation coverage and minimize unexpected issues during production.
β’ Participate in projects aimed at training and refining AI systems.
β’ Must possess native fluency in both English and Assamese.
β’ Previous experience in red teaming focused on AI adversarial tasks, cybersecurity, or socio-technical probing.
β’ Proficiency in probing AI models through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
β’ Capability to annotate failures, categorize vulnerabilities, and pinpoint systemic risks.
β’ Familiarity with following taxonomies, benchmarks, and playbooks.
β’ Competence in generating reproducible reports, datasets, and attack cases.
β’ A curious and adversarial mindset toward system testing.
β’ Ability to clearly communicate risks to both technical and non-technical stakeholders.
β’ Flexibility to adapt across various projects and clients.
β’ Status as an independent contractor is required.
β’ Candidates must not hold H1-B or STEM OPT status.
β’ Fully remote position.
β’ Flexible work schedule allowing you to choose your own hours.
β’ Weekly compensation via Stripe or Wise.
β’ Optional participation in high-sensitivity projects.
β’ Clear guidelines and wellness resources available for projects involving sensitive content.
β’ Opportunity to gain experience in human data-driven AI red teaming.
β’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Competitive compensation package.
β’ Collaboration with leading researchers in the field.
β’ Referral bonus of up to $90 for each successful referral.
Mercor
BPCS, Comprehensive marketing solutions, ltd.
Tech Mahindra
Get handpicked remote jobs straight to your inbox weekly.