Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mercor

AI Safety Expert – English, Assamese

Mercor

. Red-team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation .

Posted 9/29/2026contractRemote • United StatesMid-LevelSenior💰 $16 - $22 per hourWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in red-teaming conversational AI models and agents, with a focus on identifying vulnerabilities and generating human data through systematic analysis. Proficient in assessing AI responses for accuracy and consistency while effectively communicating findings to diverse audiences.

Highest-signal resume keywords
Red-Team Conversational AI ModelsVulnerability AssessmentAdversarial ML ExperiencePenetration TestingNative Fluency in English and Assamese

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
JailbreaksPrompt InjectionsBias ExploitationMulti-Turn ManipulationAnnotation of FailuresClassification of VulnerabilitiesSystemic Risk FlaggingReproducible ReportingEvaluation Coverage ExpansionQuality Standards Adherence
Soft Skills
Strong Judgment About LanguageClear Explanation of ReasoningAdaptability Across ProjectsAttention to DetailEffective Communication
Industry Keywords
CybersecurityExploit DevelopmentReverse EngineeringSocio-Technical RiskConversational AI TestingCreative ProbingHarassment ProbingDisinformation Probing

Tech Stack

Tools & technologies
Cyber Security

About the role

Key responsibilities & impact
  • Red-team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply taxonomies, benchmarks, and playbooks to keep testing consistent
  • Produce reproducible reports, datasets, and attack cases for customers
  • Uncover vulnerabilities that automated tests miss
  • Expand evaluation coverage and reduce production surprises
  • Work on projects focused on training and enhancing AI systems

Requirements

What you’ll need
  • Native fluency in English and Assamese
  • Strong judgment about language and content
  • Ability to assess whether AI responses are accurate, complete, and appropriate, and explain why
  • Ability to identify subtle errors, inconsistencies, and gaps
  • Ability to consistently follow guidelines and quality standards
  • Ability to explain reasoning clearly to technical and non-technical audiences
  • Adaptability across projects, task types, and customers
  • Nice-to-have: adversarial ML experience, including jailbreak datasets, prompt injection, RLHF/DPO attacks, and model extraction
  • Nice-to-have: cybersecurity experience, including penetration testing, exploit development, and reverse engineering
  • Nice-to-have: socio-technical risk experience, including harassment/disinformation probing, abuse analysis, and conversational AI testing
  • Nice-to-have: creative probing experience in psychology, acting, or writing
  • Must be engaged as an independent contractor
  • H1-B or STEM OPT candidates are not supported

Benefits

Comp & perks
  • Fully remote role
  • Flexible schedule / work on your own schedule
  • Weekly payments via Stripe or Wise
  • Projects may be extended, shortened, or concluded early depending on needs and performance
  • Higher-sensitivity projects are optional
  • Clear guidelines and wellness resources
  • Competitive pay
  • Opportunity to collaborate with leading researchers
  • Referral payments of up to $90 per successful referral (referral limits apply)