Remotery

AI Safety Experts, English – Bengali

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$20 – $22/hour

Posted 6 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team evaluations of conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

β€’ Create human data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistency in testing.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Analyze AI outputs related to sensitive issues such as bias, misinformation, and harmful behaviors.

β€’ Discover vulnerabilities that automated assessments may overlook.

β€’ Increase evaluation coverage while minimizing unexpected outcomes in production.

β€’ Enhance customer AI systems through rigorous adversarial testing.


⛳️ Requirements

β€’ Fluent or native proficiency in both English and Bengali.

β€’ Previous experience in red teaming focused on AI adversarial tasks, cybersecurity, or socio-technical probing.

β€’ Capability to challenge systems to their limits through adversarial testing.

β€’ Proficiency in utilizing frameworks or benchmarks rather than employing random hacks.

β€’ Ability to clearly communicate risks to both technical and non-technical stakeholders.

β€’ Flexibility to adapt across various projects and client needs.

β€’ Experience in adversarial machine learning is a desirable asset.

β€’ Background in cybersecurity, such as penetration testing, exploit development, or reverse engineering, is a plus.

β€’ Experience in managing socio-technical risks is considered an advantage.

β€’ Background in creative probing through psychology, acting, or writing is a nice-to-have.

β€’ Must be an independent contractor.

β€’ H1-B and STEM OPT candidates are not eligible for this position.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible scheduling options.

β€’ Weekly payments processed via Stripe or Wise.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Collaboration with leading researchers in the field.

β€’ Reasonable accommodations available upon request.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines and wellness resources provided for handling sensitive content.

β€’ Referral payments of up to $90 for each successful referral.

People also viewed

CVS Health12 hours ago

Lead Director, Digital Product – Conversational AI Strategy

US flagTexas OnlyFull-timeArtificial Intelligence$144.2k – $288.4k/year
ApplyView job
One Impression12 hours ago

AI Generalist Intern – Founder's Office

IN flagIndia OnlyInternshipArtificial Intelligence
ApplyView job
Volga Partners12 hours ago

AI Language Quality Evaluator – Greek/English, Mid-Level

GR flagGreece, +1 more countryFreelanceArtificial Intelligence$7 – $9/hour
ApplyView job
Mercor13 hours ago

Senior Design Expert – Paid AI Design Research Study

US flagUnited States OnlyFreelanceArtificial Intelligence$150 – $250/hour
ApplyView job
Reveleer13 hours ago

SVP, AI & Data

US flagUnited States OnlyFull-timeArtificial Intelligence$305k – $355k/year
ApplyView job
Mitratech13 hours ago

AI Automation Specialist

MX flagMexico OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers