Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Walmart

Systems and Infrastructure Engineer III

Walmart

. Support and operate Walmart Cloud Native Platform (WCNP), an enterprise Kubernetes platform powering cloud-native workloads .

Posted 9/15/2026full-timeChennai • IndiaMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Kubernetes administration, including cluster lifecycle management, workload diagnostics, and security controls. Proficient in operational automation using Python and Bash, with a strong focus on CI/CD pipeline management and incident response.

Highest-signal resume keywords
Kubernetes AdministrationCI/CD Pipeline ManagementPython AutomationIncident ResponseKubernetes Security

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
KubernetesCluster Lifecycle ManagementRBACAutoscalingPrometheusGrafanaHelmJenkinsGitOpsPython
Tools & Technologies
GKEAKSIstioService MeshConcordLooperServiceNow
Industry Keywords
Cloud NativeInfrastructure EngineeringPlatform OperationsDevOpsCompliance

Tech Stack

Tools & technologies
AzureCloudDNSGoogle Cloud PlatformGrafanaJenkinsKubernetesNode.jsPrometheusPythonServiceNow

About the role

Key responsibilities & impact
  • Support and operate Walmart Cloud Native Platform (WCNP), an enterprise Kubernetes platform powering cloud-native workloads
  • Serve as primary support contact for workloads across production and non-production environments
  • Manage cluster health, workload triage, incident response, and on-call operations to meet platform SLOs
  • Provision and harden clusters, execute version upgrades, rotate node pools, validate control planes, and perform maintenance
  • Build, maintain, and troubleshoot CI/CD pipelines with compliance gates and security scans
  • Troubleshoot observability issues using Prometheus, Grafana, and OpenObserve; build dashboards and alerting rules
  • Produce root-cause analyses to drive preventive action and reduce MTTR
  • Support container networking, traffic management, service mesh, DNS, TLS termination, and load balancer configurations
  • Author and manage KITT configurations, Helm-based configurations, Kubernetes manifests, RBAC, and network policies
  • Coordinate changes through ServiceNow Change Management with version control and peer review
  • Analyze workload utilization and implement right-sizing, HPA/VPA, and autoscaler strategies
  • Implement and audit Kubernetes security controls, secrets management, image scanning, and compliance measures
  • Partner with application, SRE, and security teams to onboard services and resolve platform blockers
  • Create runbooks, playbooks, and platform documentation
  • Design automation pipelines using Concord and Looper and build reusable Helm charts
  • Develop Python- and AI/LLM-based operational tools for log analysis, alert triage, and runbook automation

Requirements

What you’ll need
  • Bachelor's degree in Computer Science, Computer Engineering, Information Systems, Software Engineering, Information Technology, or related area and 5 years' experience in software engineering, platform engineering, infrastructure engineering, or a related area; OR 7 years' experience in platform engineering, infrastructure engineering, DevOps, or a related technical area
  • 6+ years of experience in infrastructure engineering, platform operations, or DevOps roles
  • Deep, hands-on experience with Kubernetes administration and operations, including cluster lifecycle management, workload diagnostics, RBAC, autoscaling, and multi-cluster management
  • Strong understanding of Kubernetes networking, Services, Ingress controllers, CoreDNS, CNI plugins, and Network Policies
  • Hands-on experience with Istio service mesh, traffic management, mTLS, and L4/L7 load balancing
  • Experience with enterprise Kubernetes platforms such as GKE or AKS
  • Experience with Helm and internal developer platform standards
  • Strong experience with Jenkins or GitHub Actions, GitOps workflows, and progressive delivery strategies
  • Proficiency with Python and Bash for operational automation and tooling development
  • Hands-on experience with Prometheus, Grafana, OpenObserve, or Datadog
  • Experience with GCP or Azure and containerised workloads
  • Working knowledge of Kubernetes security, secrets management, PCI-DSS, and SOC2
  • Basic knowledge of AI/LLM API integration and prompt engineering
  • Experience with L2/L3 incident triage, root cause analysis, post-incident reviews, and on-call operations
  • Ability to work from the Chennai office for daily work

Benefits

Comp & perks
  • Incentive awards for performance
  • Maternity and parental leave
  • PTO
  • Health benefits
  • Flexibility for associates to manage their personal lives