Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mercor

AI Safety Expert – English, Kannada

Mercor

. Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation .

Posted 9/21/2026contractRemote • United StatesMid-LevelSenior💰 $16 - $22 per hourWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in adversarial testing of conversational AI models, focusing on identifying vulnerabilities and generating human data through systematic evaluation. Proficient in assessing AI outputs for bias and misinformation while maintaining high-quality standards across diverse projects.

Highest-signal resume keywords
Adversarial TestingVulnerability AssessmentBias ExploitationNative Fluency in English and KannadaClear Communication

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
JailbreaksPrompt InjectionMisuse CasesMulti-Turn ManipulationAnnotation of FailuresClassification of VulnerabilitiesSystemic Risk FlaggingReproducible ReportingDataset ProductionEvaluation Coverage Expansion
Soft Skills
Strong JudgmentAttention to DetailAdaptabilityClear Reasoning
Industry Keywords
Conversational AICybersecurityPenetration TestingExploit DevelopmentSocio-Technical RiskDisinformation ProbingCreative Probing

Tech Stack

Tools & technologies
Cyber Security

About the role

Key responsibilities & impact
  • Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply taxonomies, benchmarks, and playbooks to maintain consistent testing
  • Produce reproducible reports, datasets, and attack cases for customers
  • Review AI outputs involving sensitive topics such as bias, misinformation, or harmful behaviors
  • Identify vulnerabilities automated tests miss
  • Expand evaluation coverage and reduce production surprises
  • Strengthen customer AI systems through adversarial testing

Requirements

What you’ll need
  • Native fluency in English and Kannada
  • Strong judgment about language and content, including assessing whether AI responses are accurate, complete, and appropriate
  • Ability to explain reasoning clearly to technical and non-technical audiences
  • Ability to notice subtle errors, inconsistencies, and gaps
  • Ability to follow guidelines and quality standards consistently
  • Adaptability across projects, task types, and customers
  • Nice-to-have: adversarial ML experience, including jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction
  • Nice-to-have: cybersecurity experience, including penetration testing, exploit development, or reverse engineering
  • Nice-to-have: socio-technical risk experience, including harassment/disinformation probing, abuse analysis, or conversational AI testing
  • Nice-to-have: creative probing experience in psychology, acting, or writing
  • Must be able to work as an independent contractor
  • H1-B and STEM OPT candidates are not supported

Benefits

Comp & perks
  • Fully remote role that can be completed on your own schedule
  • Weekly payments via Stripe or Wise based on services rendered
  • Projects may be extended, shortened, or concluded early depending on needs and performance
  • Higher-sensitivity projects are optional
  • Clear guidelines and wellness resources for higher-sensitivity projects
  • Competitive payment
  • Collaboration with leading researchers
  • Referral opportunity earning up to $90 per successful referral, subject to limits
  • Reasonable accommodations upon request