AI Safety Expert – English, Punjabi

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted Sep 19

This is a fully remote position, open to applicants in United States.

📋 Description

• Conduct red-team assessments of conversational AI models and agents through jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulations.

• Create human-generated data by annotating failures, classifying vulnerabilities, and identifying systemic risks.

• Utilize taxonomies, benchmarks, and playbooks to maintain consistency in testing.

• Generate reproducible reports, datasets, and attack cases for clients.

• Evaluate AI outputs related to sensitive issues, including bias, misinformation, or harmful behaviors.

• Assist in broadening evaluation coverage and enhancing the resilience of customer AI systems.


⛳️ Requirements

• Proficiency in both English and Punjabi is essential.

• Strong analytical skills regarding language and content; ability to evaluate the accuracy, completeness, and appropriateness of AI responses, along with the capacity to articulate the reasoning.

• Keen attention to detail to identify subtle errors, inconsistencies, and omissions.

• Consistent adherence to guidelines and quality standards.

• Capability to convey reasoning effectively to both technical and non-technical audiences.

• Flexibility to adapt to various projects, task types, and client needs.

• Must have independent contractor status.

• H1-B and STEM OPT candidates are not eligible.

• Preferred qualifications include expertise in adversarial machine learning, cybersecurity, socio-technical risk, and creative probing.


🏝️ Benefits

• Fully remote position.

• Flexible working hours; manage your own schedule.

• Weekly payments through Stripe or Wise based on services provided.

• Project timelines may be adjusted—extended, shortened, or concluded early—based on requirements and performance.

• Reasonable accommodations available upon request.

• Opportunity to gain experience in human data-driven AI red teaming.

• Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

• Access to wellness resources and clear guidelines for projects involving higher sensitivity.

• Referral bonuses of up to $90 for each successful referral.

People also viewed

WON.ai17 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner19 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers