Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Vital Tech Solutions

Senior Site Reliability Engineer, SRE

Vital Tech Solutions

. Maintain, monitor, and troubleshoot production environments to support system uptime, reliability, and performance .

Posted 9/22/2026full-timeRemote • United StatesSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates strong expertise in maintaining and troubleshooting production infrastructure, with hands-on experience in Terraform, Ansible, Docker, and AWS services such as EKS, S3, and EMR. Proficient in CI/CD pipeline support, incident response, and effective communication in collaborative environments.

Highest-signal resume keywords
TerraformAnsibleDockerAWS ServicesCI/CD Pipeline Support

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
PythonLinux Command-LineIncident ResponseRoot-Cause AnalysisMonitoring and Troubleshooting
Soft Skills
Strong Communication SkillsCollaborative Work StyleAdaptability to New Technologies
Tools & Technologies
GitSparkJupyterHubHueDatabricks
Certifications & Qualifications
Active Secret Security Clearance
Industry Keywords
Production InfrastructureOperational WorkflowsFederal Government ExperienceSecurity-Sensitive Environments

Tech Stack

Tools & technologies
AnsibleAWSCloudDockerLinuxPythonSparkTerraform

About the role

Key responsibilities & impact
  • Maintain, monitor, and troubleshoot production environments to support system uptime, reliability, and performance
  • Manage and operate infrastructure using Terraform, Ansible, and Docker
  • Support and maintain CI/CD pipelines and automate operational workflows using Git and related tooling
  • Ensure the ongoing reliability of systems operating within AWS environments, including EKS, S3, and EMR
  • Support operational use of Spark, JupyterHub, and Hue
  • Diagnose and resolve infrastructure and application issues
  • Conduct root-cause analysis and drive long-term resolution of recurring incidents
  • Implement, refine, and maintain infrastructure and application monitoring, alerting, and diagnostics
  • Support deployment activities, maintenance windows, and production changes
  • Optimize data flows and storage integrations
  • Collaborate with engineering, product, and client stakeholders to communicate issues, coordinate maintenance, and support operational priorities
  • Contribute to continuous improvement of operational processes, platform documentation, and reliability best practices

Requirements

What you’ll need
  • Active Secret security clearance or higher is required
  • Strong experience supporting and maintaining production infrastructure
  • Hands-on Python experience within operational, infrastructure, or support environments
  • Professional experience with Terraform, Ansible, and Docker
  • Experience supporting CI/CD pipelines and deployment workflows
  • Strong proficiency with Git and version-control practices
  • Strong Linux command-line and systems operations experience
  • Experience monitoring, diagnosing, and troubleshooting production systems
  • Strong incident-response and root-cause analysis capabilities
  • Ability to anticipate and resolve complex operational issues
  • Strong communication skills and the ability to work effectively in a collaborative, client-facing environment
  • Ability to learn and adapt to new technologies quickly
  • Ability to work East Coast business hours
  • Experience operating infrastructure within AWS or another major cloud platform
  • Hands-on experience with AWS services including EKS, S3, and EMR
  • Familiarity with Spark, JupyterHub, and Hue in an operational environment
  • Databricks experience
  • Experience supporting federal government, regulated, or other security-sensitive environments
  • Experience working directly with external clients or government stakeholders
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline (preferred)

Benefits

Comp & perks
  • Remote work arrangement
  • Full-time employment