AI Safety Expert, English – Urdu

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted Sep 18

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments on conversational AI models and agents utilizing jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Create human-like data by annotating failures, categorizing vulnerabilities, and identifying systemic risks.

β€’ Implement taxonomies, benchmarks, and playbooks to ensure consistency in testing.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Evaluate AI outputs related to sensitive subjects such as bias, misinformation, or harmful behaviors.

β€’ Identify vulnerabilities that automated tests may overlook.

β€’ Broaden evaluation coverage while minimizing unexpected issues during production.

β€’ Collaborate on initiatives aimed at training and improving AI systems.


⛳️ Requirements

β€’ Native fluency in both English and Urdu is essential.

β€’ Strong discernment regarding language and content.

β€’ Capability to evaluate whether AI responses are accurate, comprehensive, and appropriate, along with the ability to articulate the reasoning behind assessments.

β€’ Meticulous attention to detail regarding subtle errors, inconsistencies, and omissions.

β€’ Consistent adherence to guidelines and quality standards.

β€’ Proficiency in clearly explaining reasoning to both technical and non-technical audiences.

β€’ Flexibility to adapt across various projects, task types, and client needs.

β€’ Independent contractor status is required.

β€’ Must not need H1-B or STEM OPT sponsorship.

β€’ Preferred areas of expertise include adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Fully remote position.

β€’ Flexible working hours / the ability to work according to your own schedule.

β€’ Weekly payments processed through Stripe or Wise.

β€’ Project durations can be adjusted based on requirements and performance.

β€’ Access to wellness resources and clear protocols for projects with heightened sensitivity.

β€’ Reasonable accommodations available upon request.

β€’ Competitive compensation.

β€’ Referral program offering up to $90 for each successful referral.

People also viewed

WON.ai17 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board18 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor19 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner19 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers