FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing and enhancing Kubernetes infrastructure platforms, focusing on availability, performance, and reliability. Proficient in cloud-native tools and practices, with a strong emphasis on incident management and automation.
Highest-signal resume keywords
Kubernetes ManagementGo ProgrammingCloud-Native ToolsIncident ResponseInfrastructure Automation
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Kubernetes PrimitivesCustom Kubernetes OperatorsProduction Stability ManagementMonitoring and Incident ResolutionCloud-Native Auto-ScalingHelmGitOpsArgoCDFlux
Tools & Technologies
KarpenterCluster API
Industry Keywords
DevOpsSite ReliabilityInfrastructure Engineering
Tech Stack
Tools & technologiesCloudFluxKubernetesGo
About the role
Key responsibilities & impact- Oversee and enhance a large-scale Kubernetes infrastructure platform for availability, performance, and reliability
- Evolve internal tooling to reduce manual overhead and streamline platform delivery across global teams
- Manage Day 1 and Day 2 operations for Kubernetes clusters using modern scaling frameworks
- Partner with the engineering team to manage production environments
- Lead high-severity incident responses
- Drive platform evolution using cloud-native tools and modern infrastructure practices
Requirements
What you’ll need- Background in DevOps, Site Reliability, or Infrastructure Engineering managing production environments
- Deep working knowledge of Kubernetes primitives, including Deployments, StatefulSets, and Services
- Hands-on experience building custom Kubernetes Operators
- Strong skills in Go for systems programming and infrastructure automation
- Proven experience managing production stability, monitoring, and resolving high-severity incidents
- Experience with cloud-native auto-scaling technologies such as Karpenter or Cluster API
- Practical familiarity with Helm and GitOps tools such as ArgoCD or Flux
- Must verify identity and eligibility to work
- Criminal background check required
- Visa sponsorship is not available for this position
Benefits
Comp & perks- Corporate bonus plan
- Healthcare benefits
- Dental benefits
- Vision benefits
- Parental leave and planning
- Mental health benefits
- 401(k) plan and match
- Flex time-off
- 11 paid holidays
- Volunteer time-off
- Flexible workforce model, including fully office-based, fully remote, or hybrid workplaces
- Reasonable accommodation during the application or recruiting process
