AI Safety Expert, English, Punjabi

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Assess conversational AI models and agents by exploring vulnerabilities such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human data by annotating errors, categorizing vulnerabilities, and identifying systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing practices.

β€’ Generate reproducible reports, datasets, and attack scenarios for clients.

β€’ Evaluate AI outputs on sensitive subjects including bias, misinformation, and harmful behaviors.

β€’ Detect vulnerabilities that automated tests might overlook.

β€’ Increase evaluation coverage while minimizing unexpected outcomes in production.

β€’ Collaborate with top AI researchers on initiatives aimed at training and enhancing cutting-edge AI systems.


⛳️ Requirements

β€’ Proficiency in English and Punjabi is essential.

β€’ Strong discernment regarding language and content, with the ability to evaluate the accuracy, completeness, and appropriateness of AI responses, and articulate the reasoning behind assessments.

β€’ Keen attention to detail, with the ability to identify subtle errors, inconsistencies, and gaps.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Ability to clearly communicate reasoning to both technical and non-technical audiences.

β€’ Flexibility across various projects, task types, and client needs.

β€’ Engagement as an independent contractor is required.

β€’ Must be capable of working without access to confidential or proprietary information from other employers, clients, or institutions.

β€’ H1-B and STEM OPT candidates are not eligible.

β€’ Preferred: experience in adversarial machine learning, including jailbreak datasets, prompt injections, RLHF/DPO attacks, or model extraction.

β€’ Preferred: background in cybersecurity, such as penetration testing, exploit development, or reverse engineering.

β€’ Preferred: experience in socio-technical risk, including harassment/disinformation probing, abuse analysis, or testing of conversational AI.

β€’ Preferred: creative probing experience in fields like psychology, acting, or writing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible work schedule allowing you to manage your own time.

β€’ Weekly payment options via Stripe or Wise.

β€’ Project duration may vary based on requirements and performance.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear guidelines for content and access to wellness resources.

β€’ Reasonable accommodations available upon request.

β€’ Competitive compensation.

β€’ Earn up to $90 for each successful referral (referral limits apply).

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board17 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers