AI Safety Expert – English, Gujarati

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 20 hours ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team exercises on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation techniques.

β€’ Create human data by annotating failures, identifying vulnerabilities, and highlighting systemic risks.

β€’ Adhere to established taxonomies, benchmarks, and playbooks to ensure consistent testing.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Examine AI outputs related to sensitive subjects such as bias, misinformation, and harmful behavior.

β€’ Identify vulnerabilities that automated tests may overlook.

β€’ Enhance evaluation scope and minimize unexpected production issues.

β€’ Fortify customer AI systems through adversarial testing.


⛳️ Requirements

β€’ Native fluency in both English and Gujarati is essential.

β€’ Strong judgment regarding language and content; ability to evaluate whether AI responses are accurate, complete, and suitable, along with the capability to articulate reasoning.

β€’ Keen ability to detect subtle errors, inconsistencies, and gaps.

β€’ Consistently follow guidelines and quality standards.

β€’ Capability to communicate reasoning effectively to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Must be an independent contractor.

β€’ Candidates must not hold an H1-B or STEM OPT visa.

β€’ Preferred qualifications include expertise in adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Fully remote position that allows you to work on your own schedule.

β€’ Weekly payments via Stripe or Wise based on services provided.

β€’ Optional participation in higher-sensitivity projects.

β€’ Clear guidelines and wellness resources for projects involving sensitive content.

β€’ Competitive compensation.

β€’ Opportunity to collaborate with top researchers in the field.

β€’ Referral bonuses of up to $90 for each successful referral, subject to certain limits.

People also viewed

Cresta19 hours ago

AI Strategist

US flagUnited States OnlyFull-timeArtificial Intelligence
ApplyView job
RR Donnelley20 hours ago

AI Workflow Engineer

US flagIllinois OnlyFreelanceArtificial Intelligence$107k – $171.2k/year
ApplyView job
MaintainX20 hours ago

Senior Director, MaintainX AI

US flagCalifornia OnlyFull-timeArtificial Intelligence
ApplyView job
Motorola Solutions22 hours ago

AI Manager – Language Models

US flagArizona, +9 more statesFull-timeArtificial Intelligence$240k – $265k/year
ApplyView job
Beglaubigt.de (YC F24)23 hours ago

Operations & AI Analyst Intern

DE flagGermany OnlyInternshipArtificial Intelligence
ApplyView job
Ashby23 hours ago

AI Outbound Marketing Manager

US flagUnited States OnlyFull-timeArtificial Intelligence$120k – $165k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers