Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Cisco

Senior Site Reliability Engineer – FedRAMP

Cisco

. Deploy, operate, and maintain resilient AWS and Kubernetes-based microservices supporting the Webex for Government environment .

Posted 9/30/2026full-timeUnited StatesSenior💰 $139,300 - $203,600 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in deploying and maintaining AWS and Kubernetes-based microservices, with a strong focus on operational automation, incident response, and reliability improvements. Proficient in programming languages such as Go, Java, and Python, with a solid understanding of containerization and observability tools.

Highest-signal resume keywords
AWS DeploymentKubernetes ManagementGo ProgrammingIncident ResponseOperational Automation

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
GoJavaPythonDockerCI/CD PipelinesLinuxDistributed SystemsObservabilityProduction DebuggingMicroservice Architecture
Soft Skills
CollaborationProblem-SolvingDocumentation
Tools & Technologies
GitPrometheusGrafanaCloudWatchCloudTrailElastic StackSplunkGitLabJenkins
Certifications & Qualifications
Bachelor’s Degree in Computer ScienceMaster’s Degree in Engineering
Industry Keywords
FedRAMPGovernment CloudRegulated EnvironmentsIncident ResponseRoot-Cause Analysis

Tech Stack

Tools & technologies
AWSCloudDistributed SystemsDockerGrafanaJavaJenkinsKubernetesLinuxMicroservicesPrometheusPythonSplunkGo

About the role

Key responsibilities & impact
  • Deploy, operate, and maintain resilient AWS and Kubernetes-based microservices supporting the Webex for Government environment
  • Monitor service health and troubleshoot complex incidents in a 24x7 production environment
  • Restore service quickly during production incidents
  • Automate repeatable operational work and service provisioning
  • Improve safety, efficiency, and consistency across deployments
  • Support identity and access management, logging, hardening, vulnerability remediation, and FedRAMP continuous monitoring
  • Collaborate with engineering, security, and compliance teams to identify operational risks and improve reliability
  • Create documentation, runbooks, operational procedures, and incident reports
  • Participate in incident response, root-cause analysis, capacity planning, and continuous improvement of production services
  • Contribute to blameless post-incident reviews, corrective actions, and engineering changes

Requirements

What you’ll need
  • Bachelor’s degree in Computer Science, engineering, or a related field plus 7+ years of related experience, or equivalent practical experience, or Master’s degree plus 4 years of experience
  • 4+ years of experience in Go, Java, Python, or a comparable programming language
  • Experience building, testing, troubleshooting, and maintaining production services and operational automation
  • Experience with Git, CI/CD pipelines, Linux, distributed systems, observability, production debugging, and SRE practices
  • Experience with monitoring, incident response, reliability improvements, and safe deployment and rollback
  • Experience diagnosing, troubleshooting, and resolving highly complex production issues beyond Tier 1 and Tier 2 support
  • Experience building and maintaining containerized applications using Docker
  • Experience writing multi-stage Dockerfiles
  • Understanding of horizontally scalable microservice architecture
  • Knowledge of Kubernetes (K8s) for production containerized workloads
  • Experience in FedRAMP, government cloud, or other regulated environments preferred
  • Experience with Prometheus, Grafana, CloudWatch, CloudTrail, Elastic Stack, or Splunk preferred
  • Experience with GitLab, Jenkins, or similar CI/CD tools preferred
  • Experience with highly available, multi-region distributed systems and disaster recovery preferred
  • Ability to create runbooks, operational procedures, and audit-ready evidence
  • Experience with blameless post-incident reviews, root-cause analysis, corrective actions, and preventive engineering changes preferred

Benefits

Comp & perks
  • Medical, dental and vision insurance
  • 401(k) plan with a Cisco matching contribution
  • Paid parental leave
  • Short- and long-term disability coverage
  • Basic life insurance
  • Cisco restricted stock unit grants may be available, vesting following continued employment
  • 10 paid holidays per full calendar year
  • 1 floating holiday for non-exempt employees
  • Paid employee birthday off
  • Paid year-end holiday shutdown
  • 4 paid personal wellness days
  • 16 days of paid vacation per full calendar year for non-exempt employees
  • Flexible vacation time off program with no defined limit for eligible exempt employees
  • 80 hours of sick time off provided on hire date and each January 1st thereafter
  • Up to 80 hours of unused sick time carried forward
  • Additional paid time away for critical or emergency family issues
  • Optional 10 paid volunteer days per full calendar year
  • Annual bonuses for non-sales roles, subject to Cisco policies
  • Performance-based incentive pay for sales-plan employees