Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
FactSet

Principal Site Reliability Engineer, Kubernetes

FactSet

. Monitor, maintain, and improve the reliability and availability of production systems .

Posted 9/24/2026full-timeUnited StatesLead💰 $190,000 - $220,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates extensive experience in ensuring the reliability, scalability, and performance of systems and services, with a strong focus on Kubernetes management and automation. Proficient in collaborating with cross-functional teams to implement best practices in infrastructure and operational efficiency.

Highest-signal resume keywords
Kubernetes ManagementCloud Platform ExperienceCI/CD ToolingInfrastructure as CodeMonitoring and Observability

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
KubernetesPythonGoBashTerraformPulumiAnsiblePuppetChefGitHub Actions
Soft Skills
Problem-SolvingAnalytical SkillsCommunicationCollaborationAbility to Work Under Pressure
Tools & Technologies
PrometheusGrafanaCoralogixOpenTelemetryHelmAWSGCPAzureArgoCDHarness
Industry Keywords
Service Level ObjectivesService Level IndicatorsCapacity PlanningPerformance OptimizationIncident Response

Tech Stack

Tools & technologies
AnsibleAWSAzureChefCloudGoogle Cloud PlatformGrafanaKubernetesPrometheusPuppetPythonTerraformGo

About the role

Key responsibilities & impact
  • Monitor, maintain, and improve the reliability and availability of production systems
  • Respond to and resolve incidents, conducting thorough post-mortems to prevent recurrence
  • Define and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
  • Collaborate with development teams to build reliability into services from the ground up
  • Design and implement automation to reduce toil and improve operational efficiency
  • Participate in an on-call rotation to support critical systems
  • Contribute to capacity planning and performance optimization efforts
  • Document systems, processes, and runbooks to support the wider team
  • Ensure the reliability, scalability, and performance of systems and services
  • Work closely with development and operations teams to build and maintain robust infrastructure, automate processes, and drive engineering best practices

Requirements

What you’ll need
  • 8+ years’ experience ensuring the reliability, scalability, and performance of systems and services
  • Kubernetes (Required)
  • Hands-on experience deploying, managing, and troubleshooting workloads in Kubernetes
  • Strong understanding of core Kubernetes concepts including Pods, Deployments, Services, ConfigMaps, and Ingress
  • Experience with Kubernetes cluster management and administration
  • Familiarity with Helm for application packaging and deployment
  • Understanding of Kubernetes networking, storage, and security best practices
  • Experience with cloud platforms such as AWS, GCP, or Azure
  • Experience with CI/CD tooling such as GitHub Actions, ArgoCD, or Harness
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Coralogix, or OpenTelemetry
  • Experience with infrastructure as code tools such as Terraform or Pulumi
  • Experience with configuration management tools such as Ansible, Puppet, or Chef
  • Programming/scripting experience with Python, Go, or Bash
  • Strong problem-solving and analytical skills with a methodical approach to troubleshooting
  • Excellent communication skills with the ability to collaborate across technical and non-technical teams
  • Ability to work effectively under pressure, particularly during incident response
  • Bachelor’s degree in computer science or relevant degree