Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
EverOps

Lead Kubernetes Platform Engineer

EverOps

. Lead a two-month EKS modernization discovery .

Posted 9/29/2026full-timeRemote • United StatesSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates extensive experience in leading EKS modernization efforts, including architecture design, automated upgrades, and capacity planning. Proficient in managing production Kubernetes environments and optimizing workloads across multiple clusters while ensuring cost efficiency and performance.

Highest-signal resume keywords
Amazon EKS ExpertiseKubernetes Production ManagementAdvanced Terraform ProficiencyKarpenter ExperienceCapacity Planning and Cost Modeling

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
EKS ModernizationKubernetes OperationsTerraformPython ScriptingGo ScriptingBash ScriptingCapacity PlanningWorkload MigrationInfrastructure as CodeService Mesh
Soft Skills
LeadershipCommunicationAnalytical ThinkingCollaboration
Tools & Technologies
AWS VPCDatadogPrometheusGrafanaHelmGitOpsArgo CDFlux
Industry Keywords
DevOpsSREPlatform EngineeringInfrastructure EngineeringARM64 Migration

Tech Stack

Tools & technologies
AWSEC2FluxGrafanaJavaKubernetesNode.jsPHPPrometheusPythonTerraformGo

About the role

Key responsibilities & impact
  • Lead a two-month EKS modernization discovery
  • Baseline the multi-cluster production EKS estate, including inventory, topology, workload placement, ownership, cost, Kubernetes versions, and support status
  • Analyze failure domains, isolation boundaries, and dependency concentration
  • Design a multi-cluster target architecture with workload placement and tenancy models
  • Design and implement automated EKS upgrades, including blue/green or in-place approaches, Karpenter drift-based node rotation, add-on compatibility, and API deprecation management
  • Lead instance sizing, workload-fit analysis, capacity planning, bin-packing, network-limit analysis, and scale-up headroom planning
  • Plan and drive phased ARM64/Graviton migration across Java, Go, PHP, and Python workloads
  • Configure Karpenter NodePools, weights, overlays, and safe Spot capacity patterns
  • Model compute, support, and commitment economics and quantify savings
  • Assess and improve ingress, Gateway API, service mesh, CNI, GitOps, and infrastructure-as-code posture
  • Quantify maintenance and toil, set reduction targets, and build automation
  • Define an AI-assisted DevOps agent workstream and assess platform readiness
  • Lead the TechPod as a player-coach and serve as primary contact for customer infrastructure leadership
  • Partner with AWS specialists on capacity planning and architecture decisions
  • Produce estate baselines, architecture designs, upgrade plans, roadmaps, and executive readouts
  • Present findings and recommendations to engineering leadership

Requirements

What you’ll need
  • 8+ years in DevOps, SRE, Platform, or Infrastructure Engineering
  • 4+ years operating production Kubernetes
  • Prior technical lead, staff, or principal-level experience
  • Deep production experience with Amazon EKS at large scale
  • Hands-on ownership of EKS version upgrades across multiple production clusters
  • Advanced production experience with Karpenter (v1+)
  • Strong understanding of EC2 instance families, generations, CPU architectures, and network performance limits
  • Experience migrating production workloads to Graviton or other ARM64 platforms
  • Experience modeling compute costs and commitments, including Savings Plans, Reserved Instances, and Spot
  • Solid knowledge of AWS VPC CNI, ingress controllers, Gateway API, service mesh, and EKS load balancing
  • Advanced proficiency with Terraform
  • Experience with Helm and GitOps tooling such as Argo CD or Flux
  • Comfort using Datadog, Prometheus, Grafana, or comparable tooling
  • Strong scripting ability using Python, Go, or Bash
  • Ability to produce a current-state picture and roadmap in weeks
  • Ability to translate technical findings into cost, risk, and capacity terms
  • U.S. work authorization and sponsorship-status questions are required in the application

Benefits

Comp & perks
  • 100% Remote Workplace
  • Unlimited Paid Time Off
  • Equity
  • 401K with company contribution
  • Sponsored healthcare
  • Training and certification programs for professional growth