FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in cloud engineering and DevOps practices, with a strong focus on incident response, automation, and infrastructure management. Proficient in utilizing major cloud platforms and observability tools to enhance operational efficiency and service reliability.
Highest-signal resume keywords
Cloud EngineeringIncident ResponseInfrastructure as CodeMonitoring and AlertingScripting and Automation
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Cloud Platform ExperienceLinux AdministrationWindows Server AdministrationPython ScriptingBash ScriptingPowerShell ScriptingDockerKubernetesITIL PracticesAI-First Approach
Soft Skills
Clear CommunicationStrong TroubleshootingHigh AdaptabilityAnalytical JudgmentCalm Under Pressure
Tools & Technologies
AWSAzureGCPDatadogPagerDutyServiceNowJira Service Management
Industry Keywords
24x7 OperationsManaged ServicesChange ManagementProblem ManagementService Improvement
Tech Stack
Tools & technologiesAWSAzureCloudDockerGoogle Cloud PlatformKubernetesLinuxPythonServiceNow
About the role
Key responsibilities & impact- Operate, monitor, and continuously improve client cloud environments
- Lead response to incidents within agreed SLAs as part of a 24x7 rotation
- Monitor AWS, Azure, and/or GCP infrastructure using observability tooling; triage alerts and respond to incidents, including nights, weekends, and holidays as scheduled
- Lead incident response, root cause analysis, and problem management for production issues
- Execute and support change requests, patching, backups, and routine maintenance following change management and least-privilege access processes
- Design, build, and maintain Infrastructure as Code and CI/CD pipelines to automate provisioning and reduce manual toil
- Improve monitoring, alerting, and dashboarding to increase visibility and reduce mean time to detect and resolve
- Support security operations by applying patches, responding to vulnerability findings, and following incident escalation procedures
- Maintain runbooks, knowledge base articles, and shift handover notes for 24x7 continuity
- Collaborate with client stakeholders, account teams, and cross-functional Valtech engineering teams on service improvement initiatives
- Contribute to capacity planning, cost optimization, and reliability improvements using SLOs and SLIs
- Participate in on-call rotation and follow escalation and paging procedures for priority incidents
Requirements
What you’ll need- Advanced/Fluent English
- 5+ years of experience in cloud engineering, DevOps, SRE, or infrastructure support roles, ideally within a managed services or 24x7 operations environment
- Deep, hands-on experience with at least one major cloud platform (AWS, Azure, or GCP); multi-cloud experience is preferred
- Working knowledge of Linux and/or Windows Server administration
- Scripting/automation experience with Python, Bash, or PowerShell and familiarity with Infrastructure as Code tools
- Solid experience with containerization and orchestration using Docker and Kubernetes
- Familiarity with monitoring and incident management tooling such as Datadog, PagerDuty, ServiceNow, or Jira Service Management
- Understanding of ITIL-aligned incident, problem, and change management practices
- Strong troubleshooting skills and ability to remain calm and methodical during high-pressure incidents
- Clear written and verbal communication skills for shift handovers and client-facing updates
- Willingness and ability to work rotating shifts, including nights, weekends, and public holidays, to support 24x7 coverage
- High adaptability and curiosity around emerging tools
- Strong analytical judgment to assess AI outputs
- Systems thinking to understand full incident-to-resolution chains with AI support
- Experience with an AI-first approach to triage, RCA generation, and observability insights
Benefits
Comp & perks- Flexibility, with remote and hybrid work options (country-dependent)
- Career advancement, with international mobility and professional development programs
- Learning and development, with access to cutting-edge tools, training, and industry experts
- Competitive compensation package
