Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mercor

AI Safety Expert – English, Punjabi

Mercor

. Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation .

Posted 9/18/2026contractRemote • United StatesMid-LevelSenior💰 $16 - $22 per hourWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in red-teaming conversational AI models through adversarial testing, focusing on identifying vulnerabilities and biases while producing reproducible reports and datasets. Proficient in evaluating AI outputs for accuracy and appropriateness, with strong language skills in English and Punjabi.

Highest-signal resume keywords
Red-Team Conversational AI ModelsAdversarial TestingFluent in English and PunjabiBias and Misinformation AssessmentReport Generation and Data Annotation

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Red-TeamingAdversarial TestingData AnnotationVulnerability ClassificationBias ExploitationPrompt InjectionMulti-Turn ManipulationSystemic Risk FlaggingReproducible Report ProductionEvaluation Coverage Expansion
Soft Skills
Strong JudgmentAttention to DetailClear CommunicationAdaptabilityConsistency to Guidelines
Industry Keywords
Conversational AIAI Model EvaluationQuality StandardsHuman Data GenerationTesting Consistency

About the role

Key responsibilities & impact
  • Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply taxonomies, benchmarks, and playbooks to keep testing consistent
  • Produce reproducible reports, datasets, and attack cases
  • Probe AI outputs involving bias, misinformation, and harmful behaviors
  • Uncover vulnerabilities that automated tests miss
  • Expand evaluation coverage and reduce production surprises
  • Strengthen customer AI systems through adversarial testing

Requirements

What you’ll need
  • Fluent language skills required: English and Punjabi
  • Native fluency in English and Punjabi required
  • Strong judgment about language and content
  • Ability to assess whether AI responses are accurate, complete, and appropriate and explain why
  • Ability to notice subtle errors, inconsistencies, and gaps
  • Ability to work consistently to guidelines and quality standards
  • Ability to explain reasoning clearly to technical and non-technical audiences
  • Adaptability across projects, task types, and customers
  • Independent contractor status
  • H1-B and STEM OPT candidates are not supported

Benefits

Comp & perks
  • Fully remote role
  • Work on your own schedule
  • Payments weekly via Stripe or Wise
  • Opportunity to build experience in human data-driven AI red teaming
  • Direct role in making AI systems more robust, safe, and trustworthy
  • Competitive pay
  • Collaboration with leading researchers
  • Reasonable accommodations upon request
  • Higher-sensitivity project participation is optional
  • Clear guidelines and wellness resources for higher-sensitivity projects