FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in Site Reliability Engineering and DevOps practices, with a strong focus on automation, cloud services, and system reliability. Proficient in scripting, containerization, and configuration management, ensuring compliance with security standards and effective incident response.
Highest-signal resume keywords
Site Reliability EngineeringCloud Services (AWS, Azure)Containerization (Docker, Kubernetes)Configuration Management (Terraform)Monitoring Tools (Grafana, Datadog)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Scripting (Python, Go)Infrastructure as CodeCI/CD Tools (GitHub Actions, AWS CodeBuild)Database Technologies (RDS, Aurora, PostgreSQL)Large-Scale Distributed Systems
Soft Skills
Analytical SkillsProblem-Solving SkillsCommunication SkillsCollaboration SkillsMentoring
Tools & Technologies
DockerKubernetesTerraformGrafanaDatadog
Certifications & Qualifications
Bachelor’s Degree in Computer Science or EngineeringActive TS/SCI ClearanceEligibility for U.S. Department of State ITAR Authorizations
Industry Keywords
Site Reliability EngineeringDevOpsAutomationIncident ResponseSecurity Compliance
Tech Stack
Tools & technologiesAWSAzureCloudDistributed SystemsDockerGrafanaKubernetesMicroservicesPostgresPythonTerraformTypeScriptGo
About the role
Key responsibilities & impact- Design, implement, and maintain scalable and reliable systems
- Set up monitoring tools and create incident response plans to identify and resolve issues and implement preventative measures
- Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks
- Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions
- Work closely with development teams to enhance product reliability and streamline deployment
- Create and maintain documentation for system architecture, processes, and incident reports
- Participate in on-call rotations to provide 24/7 support for critical systems
- Implement and enforce security best practices across systems and ensure compliance with industry standards
- Independently deploy infrastructure changes using Infrastructure as Code
- Identify reliability risks and recommend improvements
- Improve dashboards, alerts, and operational runbooks
- Improve deployment pipelines, infrastructure provisioning, and self-service capabilities
- Optimize infrastructure utilization and cloud costs without compromising reliability
- Drive automation to reduce operational toil and improve deployment reliability
- Lead cross-functional initiatives to improve availability, scalability, and operational efficiency
- Advise on site reliability for new product developments and platform evolution
- Mentor junior engineers and foster the development culture
Requirements
What you’ll need- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience
- 5+ years of experience in a Site Reliability Engineering, DevOps, or related role
- Proficiency in a scripting or programming language, such as Python or Go
- Experience with cloud services such as AWS or Azure
- Proficiency with containerization using Docker, Kubernetes, or ECS
- Proficiency in configuration management tools such as Terraform, Atlantis, or Terragrunt
- Familiarity with CI/CD tools such as GitHub Actions, AWS CodeBuild, or CircleCI
- Experience with monitoring tools such as Grafana or Datadog
- Familiarity with database technologies such as RDS, Aurora, or PostgreSQL
- Experience with large-scale distributed systems and microservices architecture
- Strong analytical and problem-solving skills with the ability to troubleshoot complex systems
- Excellent verbal and written communication skills, with the ability to collaborate effectively across teams
- Ability to obtain a security clearance
- Active TS/SCI clearance preferred
- Must be eligible to obtain required U.S. Department of State ITAR authorizations
Benefits
Comp & perks- Bonus and equity
- 100 USD monthly wellness benefit to support your health and well-being
- $300 USD home office setup stipend, available for use within your first six months
- $75 USD monthly home internet allowance, paid directly through your regular paycheck
- Individual Development Fund to support training, learning, and professional development
- Employee Recognition Program to celebrate contributions and achievements across the team
- Employee Referral Program
- And much more!
