FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in building and maintaining cloud infrastructure with a focus on resilience, security, and scalability. Proficient in Infrastructure as Code, incident management, and observability practices to ensure operational excellence.
Highest-signal resume keywords
Infrastructure As CodeAWSTerraformIncident ManagementCI/CD
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Infrastructure As CodeIncident ManagementCloud ArchitectureScriptingObservabilityNetworking FundamentalsSecurity PracticesContainerizationPost-Incident ReviewsAutomated Provisioning
Soft Skills
Collaborative Team PlayerCommunication Skills
Tools & Technologies
TerraformDatadogECSPythonTypeScriptBashGCPAzure
Industry Keywords
Cloud InfrastructureIncident ResponseMonitoringLoggingSecurity Standards
Tech Stack
Tools & technologiesAWSAzureCloudGoogle Cloud PlatformPythonTerraformTypeScript
About the role
Key responsibilities & impact- Build and maintain cloud infrastructure and paved paths for resilience, security, and scalability
- Provide self-service monitoring, alerting, logging, and tracing
- Automate provisioning, deployments, and operational work
- Develop incident tooling, runbooks, and incident-response practices
- Conduct post-incident reviews and implement resulting changes
- Set identity and least-privilege models for agents, CI, and MCP servers
- Document infrastructure designs and operational procedures, including machine-readable runbooks
- Establish shipping safety standards through pipeline guardrails and checks
- Provide visibility and controls over cloud and inference spend
- Coach engineers in reliability practices
- Collaborate with Engineering and Product on feature delivery and platform needs
Requirements
What you’ll need- Experience running incident command on real production incidents
- Experience owning a platform through a meaningful scaling step
- Deep Infrastructure as Code skills with Terraform, Terragrunt, or CDK
- Experience with containers, especially ECS
- Experience with AWS preferred, or GCP/Azure
- CI/CD experience
- Scripting in Python, TypeScript, or Bash
- Observability experience with Datadog or similar
- Knowledge of distributed tracing and structured logging
- Understanding of incident management methodology
- Networking and cloud architecture fundamentals
- Understanding of security for agentic systems, including credential handling, least privilege, and prompt injection
- Experience using coding agents such as Claude Code or equivalent
- Ability to communicate system health, risks, and trade-offs to technical and non-technical audiences
- Collaborative and supportive team-player approach
Benefits
Comp & perks- Competitive market rate remuneration, reviewed twice annually
- Employee Share Option Program (ESOP)
- Annual wellness bonus
- Premium EAP platform access
- 6 weeks of paid annual leave
- 12 weeks’ paid parental leave for either caregiver
- Additional sick leave for IVF
- Gradual return to work
- $1000 personal L&D budget
- Mentorships, speaking engagements, and travel growth opportunities
- Flexible working options and WFH/in-office flexibility
- Office access in Auckland, Sydney, London, and New York
- Epic Tracksuit provided
