FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in defining SLOs, incident response practices, and building multi-region Kubernetes infrastructure. Proficient in deployment automation and hardening of PostgreSQL clusters, with a strong focus on integrating reliability and compliance into architecture.
Highest-signal resume keywords
Kubernetes ExpertiseSRE/Platform Engineering ExperienceDeployment Automation with ArgoCD and TerraformPostgreSQL Cluster Design and HardeningMonitoring with Datadog and Prometheus
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
SLO DefinitionIncident Response PracticesMulti-Region Kubernetes InfrastructureDeployment AutomationPostgreSQL ClustersRedis CachingCanary PipelinesService MeshAutoscalingSecurity Hardening
Soft Skills
CollaborationCommunication
Tools & Technologies
ArgoCDHelmTerraformDatadogPrometheusGrafana
Certifications & Qualifications
BS in Computer Science
Industry Keywords
SREPlatform EngineeringPaaSSaaSDeveloper-Facing Platforms
Tech Stack
Tools & technologiesGrafanaKubernetesPostgresPrometheusRedisTerraform
About the role
Key responsibilities & impact- Define and drive SLOs, error budgets, and incident response practices for all Volcano services
- Design and build multi-region Kubernetes infrastructure, networking, and data plane
- Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt
- Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage
- Instrument Volcano services with meaningful SLIs and build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana
- Collaborate with the OCTO team, product engineering, and security to integrate reliability and compliance into the architecture
- Evaluate architectural options for edge runtimes, serverless compute, vector databases, and AI-native infrastructure components
Requirements
What you’ll need- BS in Computer Science or equivalent
- Substantial experience at Staff or Principal IC level in SRE/Platform Engineering
- Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products
- Kubernetes expertise, including multi-tenant cluster design, networking (CNI, service mesh, ingress), autoscaling, and security hardening
Benefits
Comp & perks- Healthcare benefits
- 401(k) plan
- Short- and long-term disability benefits
- Basic life and AD&D insurance
- Additional rewards for eligible roles, including sales incentives where applicable
