FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and operating scalable, fault-tolerant systems on Microsoft Azure, with a strong focus on Infrastructure as Code, automation, and incident management. Proficient in implementing security best practices and driving continuous service improvement in cloud-native environments.
Highest-signal resume keywords
Microsoft Azure OperationsInfrastructure as Code (Bicep, ARM, Terraform)Automation and Scripting (PowerShell)Containerization (Docker, Kubernetes)Observability Tooling (Azure Monitor, Grafana, Prometheus)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Site Reliability EngineeringProduction Workload ManagementIncident Response and Root Cause AnalysisService Level Agreement (SLA) ManagementContinuous Integration/Continuous Deployment (CI/CD)Monitoring and Logging SolutionsSecurity Best PracticesProblem-Solving in Distributed SystemsOperational Performance ReportingService Health Indicators
Soft Skills
Effective CommunicationCollaboration Across TeamsStakeholder EngagementTroubleshooting CapabilityContinuous Improvement Mindset
Tools & Technologies
Azure MonitorGrafanaPrometheusDatadogOpenTelemetryBicepARMTerraformDockerKubernetes
Certifications & Qualifications
Azure Administrator AssociateAzure Solutions Architect Expert
Industry Keywords
SaaSCloud-Native EnvironmentISO27001SOC 2GDPR
Tech Stack
Tools & technologiesAzureCloudDistributed SystemsDockerGrafanaKubernetesPrometheusTerraform
About the role
Key responsibilities & impact- Design, implement, and operate highly available, scalable, and fault-tolerant systems on Microsoft Azure
- Define, track, and improve reliability metrics, service health indicators, and operational performance reporting
- Develop and maintain Infrastructure as Code using Bicep, ARM, Terraform, or similar tooling
- Build automation for provisioning, deployment, scaling, and operational workflows
- Enhance CI/CD pipelines with DevOps to improve deployment safety, reliability, and efficiency
- Implement and maintain monitoring, logging, tracing, and alerting solutions
- Define alerting strategies that reduce noise and improve response effectiveness
- Lead incident response, including troubleshooting, stakeholder communication, root cause analysis, and post-incident reviews
- Strengthen incident and problem management processes to improve SLA adherence and mitigate customer impact
- Implement systemic improvements to prevent repeat incidents
- Embed security and compliance best practices across infrastructure, including access control, encryption, and policy enforcement
- Drive continuous service improvement for performance, reliability, efficiency, and operational maturity
- Collaborate with Development and QA teams to improve application resilience and supportability
Requirements
What you’ll need- Proven experience as a Site Reliability Engineer or in a similar reliability-focused role within a SaaS or cloud-native environment
- Strong hands-on experience operating production workloads on Microsoft Azure across compute, networking, storage, and monitoring services
- Infrastructure as Code expertise using Bicep, ARM, Terraform, or similar tools
- Strong automation and scripting capability; PowerShell essential
- Experience with containerised environments (Docker) and orchestration concepts such as Kubernetes
- Practical experience with observability tooling such as Azure Monitor, Grafana, Prometheus, Datadog, or OpenTelemetry
- Strong understanding of structured incident response, root cause analysis, SLA/SLO concepts, and reliability engineering practices
- Knowledge of security best practices and compliance standards such as ISO27001, SOC 2, and GDPR
- Strong problem-solving capability with the ability to troubleshoot complex, distributed systems
- Effective communication skills and ability to collaborate across engineering, operations, and business stakeholders
- Azure certifications are desirable, including Azure Administrator Associate or Azure Solutions Architect Expert
Benefits
Comp & perks- 25 days Annual Leave
- Private pension
- Bonus scheme
- Private health
- Life assurance
