AI Safety Expert – English, Punjabi

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human data through annotating failures, identifying vulnerabilities, and highlighting systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to maintain testing consistency.

β€’ Generate replicable reports, datasets, and attack scenarios that clients can act upon.

β€’ Evaluate AI outputs related to sensitive subjects like bias, misinformation, or harmful behavior.

β€’ Discover vulnerabilities overlooked by automated testing.

β€’ Broaden evaluation scope and minimize unexpected production issues.

β€’ Enhance customer AI systems to improve their safety, robustness, and reliability.


⛳️ Requirements

β€’ Proficient fluency in both English and Punjabi is essential.

β€’ Strong discernment regarding language and content, particularly in evaluating the accuracy, completeness, and appropriateness of AI responses.

β€’ Capability to detect subtle errors, inconsistencies, and gaps.

β€’ Consistently adhere to guidelines and quality standards.

β€’ Ability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, tasks, and clients.

β€’ Independent contractor status required.

β€’ Candidates must not hold H1-B or STEM OPT status, as Mercor cannot accommodate these applicants.

β€’ Preferred: experience in adversarial machine learning, including familiarity with jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction.

β€’ Preferred: background in cybersecurity, such as penetration testing, exploit development, or reverse engineering.

β€’ Preferred: experience with socio-technical risks, including harassment/disinformation probing, abuse analysis, or testing conversational AI.

β€’ Preferred: experience in psychology, acting, or writing to foster unconventional adversarial thinking.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible working hours; tasks can be completed at your convenience.

β€’ Weekly payments through Stripe or Wise based on services provided.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Clear content guidelines and wellness resources available.

β€’ Reasonable accommodations offered upon request.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Chance to engage in cutting-edge AI safety projects.

β€’ Opportunity to collaborate with top researchers in the field.

β€’ Referral bonuses of up to $90 for each successful referral, subject to limitations.

People also viewed

WON.ai16 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor18 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner18 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers