AI Safety Expert, English, Punjabi

atMercorRemoteUS flagUnited StatesFreelanceArtificial IntelligenceMid-levelSenior$16 – $22/hour

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

πŸ“‹ Description

β€’ Conduct red-team assessments of conversational AI models and agents through techniques such as jailbreaks, prompt injections, misuse scenarios, bias exploitation, and multi-turn manipulation.

β€’ Document failures, categorize vulnerabilities, and highlight systemic risks.

β€’ Utilize taxonomies, benchmarks, and playbooks to ensure consistent testing.

β€’ Generate reproducible reports, datasets, and attack scenarios.

β€’ Evaluate AI outputs related to sensitive topics, including bias, misinformation, and harmful behaviors.

β€’ Identify vulnerabilities that automated tests might overlook.

β€’ Broaden evaluation coverage and minimize unexpected issues in production.

β€’ Enhance customer AI systems through human-driven red teaming based on data analysis.


⛳️ Requirements

β€’ Proficient fluency in both English and Punjabi is essential.

β€’ Strong discernment regarding language and content, including the ability to assess the accuracy, completeness, and appropriateness of AI responses.

β€’ Capability to articulate reasoning clearly to both technical and non-technical audiences.

β€’ Meticulous attention to detail, including subtle errors, inconsistencies, and omissions.

β€’ Consistently adhere to guidelines and quality standards.

β€’ Flexibility to work across various projects, tasks, and client needs.

β€’ Independent contractor status is required.

β€’ Candidates must not hold H1-B or STEM OPT status.

β€’ Preferred expertise includes adversarial machine learning, cybersecurity, socio-technical risk, or creative probing.


🏝️ Benefits

β€’ Fully remote position allowing you to work on your own schedule.

β€’ Weekly payments via Stripe or Wise based on services provided.

β€’ Participation in higher-sensitivity projects is optional.

β€’ Access to clear guidelines and wellness resources for handling sensitive content.

β€’ Reasonable accommodations available upon request.

β€’ Opportunity to gain experience in human data-driven AI red teaming.

β€’ Direct involvement in enhancing the robustness, safety, and trustworthiness of AI systems.

β€’ Collaborate with leading researchers in the field.

β€’ Referral bonuses of up to $90 for each successful referral, subject to limitations.

People also viewed

WON.ai15 hours ago

AI Strategist

AR flagArgentina OnlyFreelanceArtificial Intelligence
ApplyView job
The College Board16 hours ago

Director, AI Assisted Solutions

US flagUnited States OnlyFull-timeArtificial Intelligence$88k – $135k/year
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Assamese

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Experts – English, Marathi

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Mercor17 hours ago

AI Safety Expert – English, Bengali

US flagUnited States OnlyFreelanceArtificial Intelligence$16 – $22/hour
ApplyView job
Gartner17 hours ago

Director, Analyst – AI Technology Economics

GB flagUnited Kingdom OnlyFull-timeArtificial Intelligence
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers