Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mercor

AI Safety Expert – English, Gujarati

Mercor

. Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation .

Posted 9/19/2026contractRemote • United StatesMid-LevelSenior💰 $16 - $22 per hourWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in evaluating conversational AI models and agents, focusing on identifying vulnerabilities and biases while producing comprehensive reports and datasets. Proficient in applying testing benchmarks and guidelines to ensure quality and consistency across projects.

Highest-signal resume keywords
Native Fluency In EnglishNative Fluency In GujaratiEvaluation Of AI ResponsesIdentifying InconsistenciesExperience With Jailbreak Datasets

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Red-Team TestingPrompt InjectionBias ExploitationMulti-Turn ManipulationVulnerability ClassificationSystemic Risk FlaggingReproducible ReportingAutomated TestingPenetration TestingExploit Development
Soft Skills
Strong Judgment About LanguageClear Explanation Of ReasoningAdaptability Across Projects
Industry Keywords
Conversational AIJailbreaksMisuse CasesQuality StandardsHarassment ProbingDisinformation AnalysisAdversarial Thinking

About the role

Key responsibilities & impact
  • Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply taxonomies, benchmarks, and playbooks to keep testing consistent
  • Produce reproducible reports, datasets, and attack cases for customers
  • Review AI outputs involving sensitive topics such as bias, misinformation, and harmful behaviors
  • Expand evaluation coverage and uncover vulnerabilities automated tests miss

Requirements

What you’ll need
  • Native fluency in English and Gujarati is required
  • Strong judgment about language and content, including evaluating whether AI responses are accurate, complete, and appropriate
  • Ability to identify subtle errors, inconsistencies, and gaps
  • Ability to consistently follow guidelines and quality standards
  • Ability to explain reasoning clearly to technical and non-technical audiences
  • Adaptability across projects, task types, and customers
  • Independent contractor engagement
  • Must not be an H1-B or STEM OPT candidate
  • Nice-to-have: experience with jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction
  • Nice-to-have: penetration testing, exploit development, or reverse engineering
  • Nice-to-have: harassment/disinformation probing, abuse analysis, or conversational AI testing
  • Nice-to-have: psychology, acting, or writing for unconventional adversarial thinking

Benefits

Comp & perks
  • Fully remote role that can be completed on your own schedule
  • Weekly payments via Stripe or Wise based on services rendered
  • Higher-sensitivity projects are optional
  • Clear guidelines and wellness resources for sensitive-content work
  • Reasonable accommodations upon request
  • Opportunity to build experience in human data-driven AI red teaming
  • Opportunity to play a direct role in making AI systems more robust, safe, and trustworthy
  • Referral payments of up to $90 for each successful referral