FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer
Charger Logistics Inc.. Keep the production platform reliable, scalable and highly available .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in maintaining reliable and scalable production platforms, with strong capabilities in Kubernetes management, infrastructure as code, and monitoring tools. Proficient in incident response and performance tuning, particularly in cloud environments.
Highest-signal resume keywords
Kubernetes ManagementInfrastructure As CodeMonitoring And Observability ToolsCI/CD Pipeline SupportProduction Incident Handling
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
KubernetesDockerTerraformHelmPythonGoBashPostgreSQLPrometheusGrafana
Tools & Technologies
AWSAzureGCPGitHub ActionsGitLab CIJenkinsAzure DevOpsELKJaegerOpenTelemetry
Certifications & Qualifications
CKACKADAWS DevOps EngineerAzure DevOps Engineer
Industry Keywords
SREDevOpsProduction InfrastructureMicroservicesDistributed SystemsService MeshLogisticsTransportation24/7 Operations
Tech Stack
Tools & technologiesAWSAzureCloudConsulDistributed SystemsDNSDockerGoogle Cloud PlatformGrafanaJenkinsKafkaKubernetesLinuxMicroservicesPostgresPrometheusPythonRabbitMQTCP/IPTerraformGo.NET
About the role
Key responsibilities & impact- Keep the production platform reliable, scalable and highly available
- Define and track SLIs, SLOs and error budgets for critical services
- Build and improve monitoring, alerting, logging and tracing using Prometheus, Grafana, Loki, Jaeger and OpenTelemetry
- Manage and tune Kubernetes clusters and containerized workloads
- Automate repetitive operational work using Python, Go or Bash
- Build and maintain infrastructure as code with Terraform and Helm
- Support and improve CI/CD pipelines for safe, frequent deployments, including blue/green and canary rollouts
- Participate in an on-call rotation, lead incident response, troubleshoot production issues and run blameless post-mortems
- Monitor and tune PostgreSQL performance, backups and high availability with the development team
- Plan capacity, test resilience and help reduce cloud costs
- Collaborate with development, DevOps and QA teams in Canada
Requirements
What you’ll need- 6–8 years in SRE, DevOps or production infrastructure roles
- Strong hands-on experience with Kubernetes and Docker in production
- Experience with at least one major cloud platform (AWS, Azure or GCP)
- Hands-on with monitoring and observability tools (Prometheus, Grafana, ELK/Loki, Jaeger or similar)
- Infrastructure as code with Terraform, Helm or similar
- Scripting or coding skills in Python, Go or Bash
- Strong Linux and networking fundamentals (DNS, load balancing, TCP/IP, TLS)
- Experience with CI/CD tools such as GitHub Actions, GitLab CI, Jenkins or Azure DevOps
- A solid understanding of microservices and distributed systems
- Experience handling production incidents and writing root-cause analyses
- Service mesh experience (Istio, Consul or Linkerd) — nice to have
- PostgreSQL administration and performance tuning — nice to have
- Familiarity with .NET Core applications in production — nice to have
- Messaging systems such as Kafka or RabbitMQ — nice to have
- Certifications such as CKA, CKAD, or AWS/Azure DevOps Engineer — nice to have
- Experience in logistics, transportation or another 24/7 operations environment — a plus
Benefits
Comp & perks- Competitive Salary
- Healthcare Benefit Package
- Career Growth