FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Incident Management and Site Reliability Engineering (SRE), with a strong focus on minimizing service downtime and effective communication during high-pressure situations. Proficient in root cause analysis, business analysis, and collaboration across teams to drive improvements in operational processes.
Highest-signal resume keywords
Incident ManagementSite Reliability Engineering (SRE)Root Cause AnalysisAzure Cloud PlatformLinux Environment Support
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Incident ResponseRoot Cause AnalysisBusiness AnalysisMonitoring ToolsNetworking ConceptsSystem ArchitectureMean Time to Mitigate (MTTM)DataDogGrafanaCloud Platforms
Soft Skills
Strong Communication SkillsNegotiation SkillsProblem-Solving MindsetAnalytical MindsetAbility to Manage High-Pressure Situations
Tools & Technologies
DataDogGrafanaAzureGCP
Industry Keywords
Incident CommanderPost-Incident Reviews (PIR)Service DowntimeStructured HandoffsStakeholder Collaboration
Tech Stack
Tools & technologiesAzureCloudGoogle Cloud PlatformGrafanaLinux
About the role
Key responsibilities & impact- Monitor and manage P1/P2 incidents in a 16x7 operational setup
- Act as incident commander, driving bridge calls and coordinating across teams
- Ensure clear, assertive communication during incidents
- Perform root cause analysis and contribute to Post-Incident Reviews (PIR)
- Drive improvements in Mean Time to Mitigate (MTTM)
- Maintain structured handoffs between EMEA and APAC regions
- Collaborate with stakeholders and ensure accountability in resolution
- Support business analysis activities and assist in developing business requirements
- Gather and document requirements
- Analyze business processes
- Support project teams and work with senior analysts on business analysis best practices
Requirements
What you’ll need- Experience in SRE / Incident Management / Production Support
- Strong communication and negotiation skills
- Ability to manage high-pressure situations confidently
- Strong problem-solving and analytical mindset
- Ability to proactively identify risks using monitoring tools such as DataDog and Grafana dashboards
- Experience in incident response, including restarting, patching, or remediating live issues
- Strong focus on minimizing service downtime across environments
- Hands-on experience supporting on-premise Linux environments and cloud platforms, primarily Azure, with some exposure to GCP
- Solid understanding of networking concepts and system architecture
