Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Cisco

Staff Site Reliability Engineer – SRE

Cisco

. Define and drive the technical roadmap for platform reliability, scalability, and operational excellence .

Posted 9/24/2026full-timeUnited StatesLead💰 $186,900 - $267,700 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering, focusing on platform reliability, scalability, and operational excellence. Proficient in leading complex incident management and driving systemic improvements through technical leadership and collaboration.

Highest-signal resume keywords
Site Reliability EngineeringKubernetes Platform ManagementCloud Architecture (AWS, GCP)Infrastructure as Code (Terraform)CI/CD Platform Design

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Python ProgrammingGo ProgrammingCapacity PlanningPerformance EngineeringDistributed Systems DesignDeployment AutomationObservability and MonitoringIncident ManagementOperational ReadinessResiliency Reviews
Soft Skills
Technical LeadershipMentoring EngineersCollaborationRoot Cause AnalysisCommunication
Tools & Technologies
KubernetesTerraformAWSGCPMLOps
Industry Keywords
Operational ExcellenceScalabilityCloud EnvironmentsSaaS DeploymentsOn-Prem Deployments

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsGoogle Cloud PlatformKubernetesPythonTerraformGo

About the role

Key responsibilities & impact
  • Define and drive the technical roadmap for platform reliability, scalability, and operational excellence
  • Lead the architecture and evolution of deployment platforms supporting cloud and air-gapped customer environments
  • Establish reliability engineering standards, including SLOs, operational readiness, capacity planning, and resiliency reviews
  • Lead reliability and scalability initiatives across Kubernetes, deployment infrastructure, databases, and networking
  • Drive automation to eliminate operational toil and improve engineering productivity
  • Design and build internal platforms, frameworks, and tooling for reliable operations at scale
  • Lead complex production incident response and drive systemic improvements through root cause analysis and long-term remediation
  • Partner with engineering leadership on platform architecture, deployment strategy, and production readiness
  • Mentor engineers through technical leadership, design reviews, and operational guidelines
  • Collaborate with customers and internal teams to build secure, scalable, and highly reliable cloud and on-prem deployment architectures

Requirements

What you’ll need
  • 8+ years’ experience with a Bachelor’s degree, or 6+ years with a Master’s degree, or 3+ years with a PhD, or equivalent related experience
  • At least 6 years in Site Reliability Engineering, Platform, Cloud, Infrastructure Engineering, or related fields
  • 5+ years operating large-scale Kubernetes platforms in production
  • Experience designing highly available, scalable, and resilient distributed systems
  • Experience with AWS, GCP, or other public cloud platforms
  • Strong experience designing CI/CD platforms and deployment automation at scale
  • Expertise in observability, monitoring, alerting, capacity planning, and performance engineering
  • Strong programming skills in Python and/or Go
  • Deep experience with Infrastructure as Code, such as Terraform
  • Experience with MLOps preferred
  • Strong understanding of networking, distributed systems, storage, databases, and cloud architecture
  • Experience operating both SaaS and enterprise/on-prem deployments
  • Demonstrated technical leadership across multiple engineering teams
  • Experience leading incident management, postmortems, and long-term reliability initiatives

Benefits

Comp & perks
  • Medical, dental and vision insurance
  • 401(k) plan with a Cisco matching contribution
  • Paid parental leave
  • Short- and long-term disability coverage
  • Basic life insurance
  • Cisco restricted stock unit grants may be available
  • 10 paid holidays per full calendar year
  • 1 floating holiday for non-exempt employees
  • Paid birthday day off
  • Paid year-end holiday shutdown
  • 4 paid personal wellness days
  • 16 days of paid vacation for non-exempt employees
  • Flexible vacation time off program for exempt employees
  • 80 hours of sick time off provided on hire date and each January 1st
  • Up to 80 hours of unused sick time carried forward
  • Additional paid time away for critical or emergency family issues
  • Optional 10 paid volunteer days per full calendar year
  • Annual bonuses for non-sales roles, subject to Cisco’s policies
  • Incentive compensation for sales roles, subject to applicable Cisco plans