FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Staff Site Reliability Engineer
Lyrebird Health. Own the reliability, security, and scalability of Lyrebird's production platform .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing AWS infrastructure, deploying CI/CD pipelines, and implementing observability strategies to ensure platform reliability and security. Proven ability to influence technical direction and improve operational maturity across engineering teams.
Highest-signal resume keywords
AWS Infrastructure ManagementTerraform AutomationCI/CD Pipeline DevelopmentIncident Response LeadershipObservability Design
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Production Workload ManagementNetworkingSecurityProduction TroubleshootingCoding for Automation
Soft Skills
Influencing Technical DirectionCross-Team Collaboration
Tools & Technologies
ECS/FargatePostgreSQLTypeScriptNode.jsOpenTelemetry
Certifications & Qualifications
ISO 27001Cyber Essentials Plus
Industry Keywords
Healthcare TechnologyDisaster RecoveryOperational MaturitySLO DefinitionIncident Detection
Tech Stack
Tools & technologiesAWSJavaScriptNode.jsPostgresTerraformTypeScript
About the role
Key responsibilities & impact- Own the reliability, security, and scalability of Lyrebird's production platform
- Manage AWS infrastructure, deployment pipelines, observability, and incident detection and response
- Define SLOs, alerting, and observability strategy
- Hold the line on platform availability and latency for clinicians
- Shape incident detection, response, and learning processes
- Evolve AWS and Terraform infrastructure for resilience, scale, and disaster recovery
- Set direction for CI/CD pipelines and developer tooling to make deployments safer and faster
- Build security controls and compliance automation into the platform
- Raise operational maturity across engineering teams so they can run their own services confidently
- Challenge technical decisions when reliability or security is at stake
- Report to the Engineering Manager, Platform and Security, and work across every engineering team
Requirements
What you’ll need- Deep hands-on experience running production workloads on AWS with Terraform, containers, and CI/CD
- Strong grounding in networking, security, and production troubleshooting
- Track record of designing observability, defining SLOs, and leading incident response that measurably improved reliability
- Coding ability to automate operational work
- Evidence of influencing technical direction across multiple engineering teams without relying on authority
- Nice to have: Experience with ECS/Fargate, PostgreSQL, and distributed or asynchronous workloads
- Nice to have: Familiarity with TypeScript/Node.js and OpenTelemetry
- Nice to have: Exposure to healthcare technology or frameworks such as ISO 27001 or Cyber Essentials Plus
- Nice to have: Experience operating platforms across multiple regions
- Right to work in Australia on an ongoing basis without requiring visa sponsorship now or in the future
Benefits
Comp & perks- Inclusive, safe, supportive, and thriving workplace culture
- Encouragement for applicants from underrepresented backgrounds in tech