Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mercor

AI Safety Experts, English, Kannada

Mercor

. Red-team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation .

Posted 9/29/2026contractRemote • United StatesMid-LevelSenior💰 $16 - $22 per hourWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in adversarial machine learning, including experience with jailbreak datasets and prompt injection, while effectively evaluating AI outputs for bias and misinformation. Capable of producing detailed reports and datasets, and communicating findings clearly to diverse audiences.

Highest-signal resume keywords
Adversarial Machine LearningCybersecurity ExperienceFluency in English and KannadaEvaluation of AI OutputsCreative Probing

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
JailbreaksPrompt InjectionsBias ExploitationMulti-Turn ManipulationVulnerability ClassificationAdversarial TestingPenetration TestingExploit DevelopmentReverse EngineeringHarassment/Disinformation Probing
Soft Skills
Strong JudgmentClear CommunicationAttention to DetailAdaptabilityConsistency in Quality Standards
Tools & Technologies
TaxonomiesBenchmarksPlaybooks
Industry Keywords
Conversational AISystemic RisksHuman Data AnnotationAI VulnerabilitiesSocio-Technical Risk

Tech Stack

Tools & technologies
Cyber Security

About the role

Key responsibilities & impact
  • Red-team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply taxonomies, benchmarks, and playbooks to keep testing consistent
  • Produce reproducible reports, datasets, and attack cases for customers
  • Review AI outputs involving sensitive topics such as bias, misinformation, and harmful behaviors
  • Uncover vulnerabilities that automated tests miss
  • Expand evaluation coverage and reduce production surprises
  • Strengthen customer AI systems through adversarial testing

Requirements

What you’ll need
  • Native fluency in English and Kannada
  • Strong judgment about language and content, including evaluating whether AI responses are accurate, complete, and appropriate
  • Ability to explain reasoning clearly to technical and non-technical audiences
  • Ability to notice subtle errors, inconsistencies, and gaps
  • Ability to follow guidelines and quality standards consistently
  • Ability to adapt across projects, task types, and customers
  • Independent contractor status
  • Must not be an H1-B or STEM OPT candidate
  • Nice-to-have specialties: adversarial ML, cybersecurity, socio-technical risk, or creative probing
  • Adversarial ML experience with jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction is a plus
  • Cybersecurity experience with penetration testing, exploit development, or reverse engineering is a plus
  • Socio-technical risk experience with harassment/disinformation probing, abuse analysis, or conversational AI testing is a plus
  • Creative probing experience in psychology, acting, or writing is a plus

Benefits

Comp & perks
  • Fully remote role
  • Flexible schedule; work can be completed on your own schedule
  • Weekly payments via Stripe or Wise
  • Projects may be extended, shortened, or concluded early depending on needs and performance
  • Participation in higher-sensitivity projects is optional
  • Clear guidelines and wellness resources for sensitive content
  • Reasonable accommodations upon request
  • Competitive payment
  • Opportunity to collaborate with leading researchers
  • Referral payments of up to $90 per successful referral, subject to limits