FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior/Lead Software Engineer, Site Reliability
Salesforce. Own the reliability roadmap for major product areas and evolve architectures into highly available, globally scalable systems .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in building and maintaining highly available, scalable systems with a strong focus on infrastructure-as-code, deployment security, and backend engineering. Proficient in leveraging AI tools for operational efficiency and mentoring engineering teams in complex distributed environments.
Highest-signal resume keywords
KubernetesTerraform/OpenTofuAWS/GCP/AzureGolangDistributed Systems
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Production-Level Code ReviewData ModelingDatabase Performance TuningAPI DesignConcurrencyFault ToleranceScalingMicroservice OrchestrationIdentity ManagementCapacity Planning
Soft Skills
Excellent WritingTechnical-Design Discussion SkillsMentoring
Tools & Technologies
AI Development ToolsCI/CD SystemsJenkinsSpinnakerRedis/ElastiCacheRDSEKSTemporalIstioSalesforce
Certifications & Qualifications
Advanced Degree in Computer ScienceM.S. Preferred
Industry Keywords
Regulated IndustriesPublic SectorData SovereigntySupply ChainLogisticsManufacturingOperational WorkflowsDepartment of Defense Impact LevelsSecurity ComplianceLimited-Connectivity Environments
Tech Stack
Tools & technologiesAWSAzureCloudDistributed SystemsGoogle Cloud PlatformGraphQLJenkinsKubernetesMicroservicesPythonRedisSpinnakerTerraformTypeScriptGo
About the role
Key responsibilities & impact- Own the reliability roadmap for major product areas and evolve architectures into highly available, globally scalable systems
- Partner with engineers on infrastructure strategy, system design, capacity planning, bottleneck identification, and security
- Define, maintain, and evolve infrastructure-as-code and deployment tools
- Support scaling and deployment of AI/ML infrastructure
- Design and build systems and processes for product testing and releases
- Build tooling and automation for customer and partner deployment, upgrades, diagnosis, and operations
- Design, review, and strengthen identity and access controls, networking, workload isolation, secret management, deployment security, and policy enforcement
- Use AI tools to automate operational tasks and accelerate infrastructure delivery
- Diagnose and prioritize architectural problems across distributed backend systems
- Design and implement application performance, infrastructure, and data-model optimizations
- Establish engineering standards, guardrails, and team practices to prevent regressions
- Collaborate with customer, product, and engineering teams to understand needs and make pragmatic tradeoffs
- Mentor engineers and share expertise in backend architecture and system design
Requirements
What you’ll need- Experience supporting mission-critical production systems and scaling high-growth products
- Strong proficiency in Kubernetes, Terraform/OpenTofu, and AWS/GCP/Azure
- Ability to write and review production-level code in Golang, TypeScript, or Python
- Deep understanding of distributed systems and debugging interactions between microservices, databases, and AI agents
- Experience working within a senior team of Principal engineers
- Demonstrated ability to use modern AI development tools
- 5+ years of experience in SRE, Production Engineering, or Backend Engineering with a heavy focus on operations and infrastructure
- Mastery of Golang, GraphQL, and PSQL
- Experience with RDS, Redis/ElastiCache, and EKS
- Expertise in backend engineering including data modeling, database performance tuning, state management, API design, transactionality, concurrency, memory management, fault tolerance, and scaling
- Repeated ownership of foundational system work in complex distributed systems
- Excellent writing, speaking, and technical-design discussion skills
- Advanced Degree in Computer Science or equivalent practical experience is an advantage
- Experience in regulated industries, particularly public sector or environments with strong security, compliance, and data-sovereignty requirements is an advantage
- Familiarity with classified or limited-connectivity environments, including Department of Defense Impact Levels such as IL6, is an advantage
- Advanced knowledge of microservice orchestration and durability patterns, including Temporal and service mesh, is an advantage
- Deep knowledge of networking, security, and identity management within major cloud providers is an advantage
- Experience building AI products for supply chain, logistics, manufacturing, or operational workflows is an advantage
- B.S. in Computer Science; M.S. preferred
- Exposure to Temporal, Istio, or Typesense
- Experience with CI/CD systems, specifically Jenkins and Spinnaker
- Exposure to supply chain, logistics, or manufacturing industry
- Familiarity with the Salesforce platform
- Experience with workflow engines
Benefits
Comp & perks- Time off programs
- Medical insurance
- Dental insurance
- Vision insurance
- Mental health support
- Paid parental leave
- Life insurance
- Disability insurance
- 401(k)
- Employee stock purchasing program
- Certain roles may be eligible for incentive compensation
- Certain roles may be eligible for equity
- Reasonable accommodations during the application or recruiting process