FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Site Reliability Engineer, GovCloud 24/7
Salesforce. Maintain system reliability and high performance across customer-facing Salesforce GovCloud services .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in maintaining high-performance Salesforce GovCloud services while ensuring compliance with security standards. Proficient in incident management, automation of operational processes, and collaboration across cross-functional teams.
Highest-signal resume keywords
Salesforce GovCloud ServicesIncident ManagementAWS/C2S InfrastructureTCP/IP TechnologiesLinux Administration
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
PythonGoUnix AdministrationJenkinsKubernetesChefPuppetMonitoring SystemsAgile ProcessesDevOps
Soft Skills
Effective CommunicationMentoringCollaborationProblem ResolutionContinuous Learning
Tools & Technologies
AWS CLIAWS SDKsSecurity Monitoring SystemsIncident Analysis ToolsAI Tools
Certifications & Qualifications
Linux+Red Hat CertificationAWS Certification
Industry Keywords
Site Reliability EngineeringOperational GuidelinesRoot Cause AnalysisHigh-Availability EnvironmentBlameless Retrospectives
Tech Stack
Tools & technologiesAWSChefJavaJenkinsKubernetesLinuxPuppetPythonSpinnakerTCP/IPUnixGo
About the role
Key responsibilities & impact- Maintain system reliability and high performance across customer-facing Salesforce GovCloud services
- Lead incident response during major operational events, including Sev0/Sev1 incidents
- Participate in post-incident reviews and drive problem resolution
- Populate and participate in Root Cause Analyses (RCAs)
- Partner with Global Solutions teams to implement preventative solutions
- Align Site Reliability Engineering activities with compliance, security standards, and operational guidelines
- Automate detection and resolution of recurring production issues
- Improve workflows to streamline engineering practices and minimize operational and engineering toil
- Collaborate with cross-functional teams on complex technical challenges
- Mentor and collaborate with teammates on emerging technologies and continuous learning
- Support dynamic operational needs and prioritize competing tasks
Requirements
What you’ll need- U.S. citizen (U.S. born or naturalized) without dual citizenship
- Successful background investigation and ability to obtain and maintain a Minimum Background Investigation (MBI) for a Moderate Public Trust position or other applicable clearance
- Must operate on U.S. soil and meet customer and government screening standards, including Criminal Justice Information Services screening with fingerprint scan
- Related technical degree
- Experience with enterprise-scale internet service operations or systems engineering
- Expertise in TCP/IP-related technologies
- Command-line expertise and Unix administration knowledge, including Linux, Solaris, or BSD
- Comprehension of security monitoring systems and infrastructure administration
- Effective written and verbal communication skills
- Familiarity with Incident Management principles and IT service operation frameworks
- Ability to support a 24/7 high-availability operational environment, including shift and on-call rotations
- Experience provisioning, operating, and running AWS/C2S infrastructure and systems
- Proficiency in Python, Go, or similar scripting/programming languages
- Prior Chef/Puppet or automated deployment experience
- Jenkins/Bamboo/Spinnaker pipeline execution experience
- Experience maintaining monitoring and alert systems
- Experience supporting and maintaining Java applications
- Hands-on AWS configuration and operation using CLI/SDKs
- Linux+, Red Hat, or AWS certifications
- Experience supporting and leading Kubernetes-based applications and services
- Experience with blameless retrospectives, incident analysis, post-incident investigations, and responder performance evaluations
- Working knowledge of resilience engineering and Safety II concepts
- Familiarity with Agile and DevOps processes
- Experience using AI tools such as Claude Code, GitHub Copilot, Codex, or Cursor
- Advanced prompt engineering skills and ability to create precise, structured prompts and reliable system context
Benefits
Comp & perks- Rotating 24/7 shift schedule with compensation for shift differentials
- Time off programs
- Medical insurance
- Dental insurance
- Vision insurance
- Mental health support
- Paid parental leave
- Life insurance
- Disability insurance
- 401(k)
- Employee stock purchasing program
- Job shadowing
- Mentorship programs
- Talent development courses