
AI Safety Expert, English, Finnish
Posted 10 hours ago

Posted 10 hours ago
This is a fully remote position, open to applicants in United States.
• Conduct red team assessments on conversational AI models and agents through methods such as jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
• Document failures, categorize vulnerabilities, and highlight systemic risks.
• Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing procedures.
• Generate reproducible reports, datasets, and attack scenarios.
• Investigate sensitive AI outputs related to bias, misinformation, and harmful behaviors.
• Broaden evaluation scope and uncover vulnerabilities that automated tests may overlook.
• Work in collaboration with leading researchers to aid in the training and enhancement of advanced AI systems.
• Proficient in both English and Finnish.
• Previous experience in red teaming focused on AI adversarial efforts, cybersecurity, or socio-technical probing.
• Capability to adversarially probe AI systems and challenge them to their limits.
• Familiarity with frameworks or benchmarks for systematic testing.
• Competence in clearly communicating risks to both technical and non-technical audiences.
• Flexibility to adapt to various projects and clientele.
• Must have independent contractor status.
• Must not require H1-B or STEM OPT sponsorship.
• Desirable areas of expertise include adversarial ML, cybersecurity, socio-technical risk, or innovative probing techniques.
• Fully remote working arrangement.
• Flexible schedule tailored to your needs.
• Weekly payments processed through Stripe or Wise.
• Competitive compensation.
• Reasonable accommodations available upon request.
• Optional involvement in high-sensitivity projects.
• Clear guidelines and wellness resources provided for high-sensitivity projects.
• Referral bonuses of up to $250 for each successful referral.
Humana
Aviva
Mercor
Get handpicked remote jobs straight to your inbox weekly.