
AI Safety Expert β English, Marathi
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United States.
β’ Engage in red-teaming of conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
β’ Generate human data by annotating failures, classifying vulnerabilities, and identifying systemic risks.
β’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.
β’ Create reproducible reports, datasets, and attack scenarios.
β’ Evaluate AI outputs concerning sensitive issues like bias, misinformation, and harmful behaviors.
β’ Identify vulnerabilities that automated testing might overlook.
β’ Enhance evaluation coverage and fortify customer AI systems.
β’ Native proficiency in both English and Marathi.
β’ Strong judgment regarding language and content accuracy, completeness, and appropriateness.
β’ Capability to identify subtle errors, inconsistencies, and omissions.
β’ Consistent adherence to guidelines and quality standards.
β’ Ability to articulate reasoning clearly to both technical and non-technical audiences.
β’ Flexibility to adapt across various projects, task types, and clientele.
β’ Must have independent contractor status.
β’ H1-B and STEM OPT candidates are not eligible.
β’ Preferred specialties include adversarial ML, cybersecurity, socio-technical risk, or creative probing.
β’ Fully remote position.
β’ Flexible scheduling to suit your needs.
β’ Weekly payments via Stripe or Wise.
β’ Opportunity to gain experience in human data-driven AI red teaming.
β’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.
β’ Collaboration with prominent researchers in the field.
β’ Reasonable accommodations available upon request.
β’ Referral bonuses of up to $90 for each successful referral.
The College Board
Mercor
Mercor
Get handpicked remote jobs straight to your inbox weekly.