FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in service management practices, including incident response and operational automation, while effectively communicating and presenting technical content. Proficient in using monitoring and incident management tools to enhance system reliability and performance.
Highest-signal resume keywords
Incident ResponseService ManagementDatadogPublic SpeakingScripting Languages
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Platform EngineeringSite Reliability EngineeringDevOps EngineeringSoftware DevelopmentIncident ManagementAPIsIaaS Cloud ServicesContainersScriptingProgramming
Soft Skills
Effective CommunicationCoachingStorytellingContent CreationSelf-Driven Learning
Tools & Technologies
DatadogPagerDutyOpsgenieIncident.ioRootlyJira Cloud PlatformCortexServiceNowJira Service Management
Industry Keywords
DevOpsMonitoringObservabilitySecurityITSMSREIncident-Response Communities
Tech Stack
Tools & technologiesCloudITSMJavaScriptNode.jsOpen SourcePythonServiceNowGo
About the role
Key responsibilities & impact- Act as a subject matter expert for service management, including incident response, on-call, IDP, Work Management, Workflow Automation, Agent Builder, and operational automation
- Create content through demos, public speaking, blogging, documentation, webinars, open source, research reports, and other mediums
- Build Datadog's reputation as a leader in DevOps, Monitoring, Observability, and Security
- Partner with product engineering teams to build compelling demos
- Coach internal engineering teams on effective communication and presentation
- Interface with open source communities to drive key messaging and develop new Datadog integrations
- Contribute to the product through feedback, documentation, or code
Requirements
What you’ll need- Approximately 5+ years of experience as a Platform Engineer, Site Reliability Engineer, DevOps Engineer or Software Developer, with hands-on experience as an on-call/incident responder and running production systems in complex IT environments
- Strong understanding of core service-management practices, including incident response, on-call, post incident reviews, and SLOs
- Experience using tools such as Datadog, PagerDuty, Opsgenie, incident.io, Rootly, Jira Cloud Platform, Cortex, or similar
- Approximately 2+ years of experience in storytelling and creating compelling content
- Publicly available writing samples, blog posts, demos, or recordings of presentations on technical topics
- Comfortable with at least one scripting and programming language, such as Node.js, Python, Go, or bash
- Familiarity with APIs, IaaS cloud services, and containers
- Self-driven exploration and education on new technologies and languages
- Hands-on ITSM platform experience with ServiceNow or Jira Service Management is a bonus
- Active membership in SRE, DevOps, or incident-response communities is a bonus
Benefits
Comp & perks- High income earning opportunities based on individual performance
- New hire stock equity (RSUs) and employee stock purchase plan (ESPP)
- Continuous professional development, product training, and career pathing
- Sales training in MEDDIC and Command of the Message
- Intradepartmental mentor and buddy program for in-house networking
- An inclusive company culture, ability to join our Community Guilds (Datadog employee resource groups)
- Access to Inclusion Talks, our Internal panel discussions
- Free, global Spring Health benefits for employees and dependents age 6+
- Competitive global benefits
- Giving programs
