
AI Safety Expert, English β Punjabi
Posted 12 hours ago

Posted 12 hours ago
This is a fully remote position, open to applicants in United States.
β’ Engage in red-teaming for conversational AI models and agents
β’ Assess vulnerabilities through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
β’ Document AI failures and categorize vulnerabilities
β’ Identify and report systemic risks
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing methodologies
β’ Generate reproducible reports, datasets, and attack scenarios
β’ Investigate AI systems for vulnerabilities that automated testing may overlook
β’ Enhance evaluation coverage and mitigate unexpected production issues
β’ Foster customer confidence in the safety of AI technologies
β’ Proficiency in both English and Punjabi is essential
β’ Previous experience in red teaming within AI adversarial contexts, cybersecurity, or socio-technical probing is necessary
β’ Capability to conduct adversarial probing of AI systems, covering jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
β’ Experience in producing human data through the annotation of failures, classification of vulnerabilities, and identification of systemic risks
β’ Competence in adhering to established taxonomies, benchmarks, and playbooks
β’ Ability to create reproducible reports, datasets, and attack scenarios
β’ Skill in articulating risks to both technical and non-technical audiences
β’ Flexibility to adapt across varying projects and client requirements
β’ Status as an independent contractor is required
β’ Candidates must not hold H1-B or STEM OPT status
β’ Position is fully remote
β’ Flexible working hours, allowing you to manage your own schedule
β’ Weekly payments facilitated through Stripe or Wise
β’ Project timelines may be adjusted based on needs and performance
β’ Access to wellness resources and clear protocols for sensitive projects
β’ Reasonable accommodations available upon request
β’ Chance to gain experience in human data-driven AI red teaming
β’ Opportunity to directly contribute to enhancing the robustness, safety, and trustworthiness of AI systems
β’ Competitive compensation
β’ Collaboration with leading researchers in the field
Mercor
Riva Scientific
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.