Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
OpenAI

Model Policy Manager, Agentic Safety

OpenAI

. Identify vulnerabilities emerging as models interact with tools, data, and external systems .

Posted 9/15/2026full-timeSan Francisco • California • United StatesMid-LevelSenior💰 $207,000 - $295,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in AI agent safety, cybersecurity, and the development of safety policies and safeguards. Capable of translating complex alignment risks into measurable evaluation criteria while collaborating across research, engineering, and product teams.

Highest-signal resume keywords
AI Agent SafetyCybersecurityEmpirical Evidence DevelopmentTechnical Fluency in Evaluation DataClear Communication of Technical Risks

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Vulnerability IdentificationThreat ModelingData Quality AssessmentPolicy Framework DevelopmentBehavioral Expectation Translation
Soft Skills
Adversarial MindsetCollaboration Across TeamsAnalytical Thinking
Industry Keywords
AI AlignmentModel BehaviorSafety PoliciesEmerging RisksSystem Safety

Tech Stack

Tools & technologies
Cyber Security

About the role

Key responsibilities & impact
  • Identify vulnerabilities emerging as models interact with tools, data, and external systems
  • Translate vulnerabilities into model- and system-level safeguards
  • Develop threat models and empirical frameworks for harmful outcomes from misaligned behavior
  • Identify underlying behaviors and system conditions driving harmful outcomes
  • Turn findings into policy frameworks, evaluation criteria, online measurement, and safeguards
  • Develop human data campaigns and gold sets for measuring emerging behaviors and risks
  • Partner with research, engineering, security, and product teams on model and system safety
  • Balance trade-offs between safety, utility, and business risk
  • Inform deployment decisions, system cards, safeguards reports, and OpenAI’s approach to agentic safety
  • Build monitoring approaches to detect post-deployment regressions and emerging risks

Requirements

What you’ll need
  • Strong background in AI agent safety, privacy, security, cybersecurity, or adjacent fields
  • Adversarial mindset for investigating real-world harmful outcomes
  • Demonstrated interest in AI alignment
  • Strong understanding of the technical drivers of misaligned model behavior
  • Technical fluency to work directly with evaluation and training data
  • Ability to inspect examples, analyze failure patterns, assess data quality, and distinguish policy failures from grader, model, or system failures
  • Experience using empirical evidence to develop and refine safety policies and safeguards
  • Ability to translate ambiguous alignment risks into precise behavioral expectations and measurable evaluation criteria
  • Ability to work across research, engineering, security, product, and policy
  • Clear communication about complex and uncertain technical risks

Benefits

Comp & perks
  • Equity
  • Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts
  • Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)
  • 401(k) retirement plan with employer match
  • Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents)
  • Paid medical and caregiver leave (up to 8 weeks)
  • Flexible PTO for exempt employees and up to 15 days annually for non-exempt employees
  • 13+ paid company holidays
  • Multiple paid coordinated company office closures throughout the year for focus and recharge
  • Paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)
  • Mental health and wellness support
  • Employer-paid basic life and disability coverage
  • Annual learning and development stipend
  • Daily meals in offices
  • Meal delivery credits as eligible
  • Relocation support for eligible employees
  • Charitable donation matching may be provided
  • Wellness stipends may be provided
  • Hybrid model with optional work from home on Thursdays and Fridays
  • Height-adjustable desks, conference rooms, phone booths, well-stocked kitchens with snacks and drinks, private outdoor space, nap rooms, and private bike storage