Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Accela

Site Reliability Engineer

Accela

. Monitor and maintain production cloud environments for availability, performance, scalability, and reliability .

Posted 10/6/2026full-timeRemote • United StatesMid-LevelSenior💰 $130,000 - $150,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in monitoring and maintaining production cloud environments, with a strong focus on Azure services, incident response, and operational efficiency. Proficient in implementing observability solutions and collaborating across teams to ensure service reliability and performance.

Highest-signal resume keywords
Site Reliability Engineering (SRE)Microsoft Azure Cloud ServicesAzure Kubernetes Service (AKS)Datadog Monitoring and LoggingIncident Response Leadership

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Cloud OperationsInfrastructure OperationsProduction Cloud EnvironmentsRoot Cause AnalysisOperational RunbooksMonitoring SolutionsTroubleshootingData MovementPerformance MonitoringAutomation Implementation
Soft Skills
Analytical SkillsProblem-Solving SkillsWritten CommunicationVerbal CommunicationCollaboration
Tools & Technologies
Azure SQLAzure StorageAzure Front DoorDatadogSaaS Platforms
Certifications & Qualifications
Bachelor's Degree in Computer ScienceBachelor's Degree in Computer Engineering
Industry Keywords
Cloud MigrationsProduction ReleasesService DegradationOperational DocumentationCustomer Provisioning

Tech Stack

Tools & technologies
AzureCloudKubernetesSQL

About the role

Key responsibilities & impact
  • Monitor and maintain production cloud environments for availability, performance, scalability, and reliability
  • Build, configure, and optimize Datadog monitoring, logging, and alerting capabilities
  • Develop and maintain dashboards for platform health, resource utilization, and service performance
  • Configure and tune alerts to identify service degradation and operational issues
  • Support and optimize Azure infrastructure, including AKS, Azure SQL, Azure Storage, and Azure Front Door
  • Execute production releases, operational changes, and platform maintenance
  • Participate in and lead incident response activities
  • Diagnose and resolve production incidents across infrastructure, platform, and application components
  • Conduct root cause analysis and implement corrective and preventative actions
  • Maintain operational documentation, runbooks, and incident response procedures
  • Implement automation to improve operational efficiency and reliability
  • Provide Level 3 support for customer-reported incidents and service requests
  • Collaborate with Engineering, Professional Services, and Operations teams on complex technical issues
  • Meet SLAs for ticket response and resolution
  • Support customer provisioning, environment maintenance, platform upgrades, and cloud migrations
  • Support production-to-non-production data refreshes
  • Assist with database operations, troubleshooting, and data movement
  • Validate operational activities and maintain documentation and records
  • Support critical production activities outside normal business hours when required, including customer go-lives and cloud migrations

Requirements

What you’ll need
  • Bachelor's degree in Computer Science, Computer Engineering, or a related technical field
  • Minimum 3 years of experience in Site Reliability Engineering (SRE), Cloud Operations, DevOps, Infrastructure Operations, or a related role
  • Minimum 3 years of hands-on experience supporting production cloud environments
  • Experience administering and supporting Microsoft Azure cloud services
  • Hands-on experience with Azure Kubernetes Service (AKS), Azure SQL, Azure Storage, and Azure Front Door
  • Experience implementing and supporting monitoring, logging, and observability solutions using Datadog
  • Experience leading incident response activities, performing root cause analysis, and developing operational runbooks
  • Strong understanding of infrastructure and application performance monitoring, including compute, memory, storage, networking, and availability metrics
  • Experience supporting SaaS platforms in production environments
  • Strong troubleshooting, analytical, and problem-solving skills
  • Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams

Benefits

Comp & perks
  • Annual bonus target (discretionary, based on company and individual goal achievement)
  • Flexible time off
  • Comprehensive medical, dental, and vision plans
  • Family planning benefits
  • 401(k) retirement savings plan with company match
  • Health savings account with company contributions
  • Flexible spending account
  • Life, accident, and disability coverage
  • Business travel insurance
  • Employee assistance programs
  • Other well-being benefits
  • Job accommodations available upon request