Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Lyrebird Health

Staff Site Reliability Engineer

Lyrebird Health

. Own the reliability, security, and scalability of Lyrebird's production platform .

Posted 10/2/2026full-timeMelbourne • AustraliaLeadWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in managing AWS infrastructure, deploying CI/CD pipelines, and implementing observability strategies to ensure platform reliability and security. Proven ability to influence technical direction and improve operational maturity across engineering teams.

Highest-signal resume keywords
AWS Infrastructure ManagementTerraform AutomationCI/CD Pipeline DevelopmentIncident Response LeadershipObservability Design

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Production Workload ManagementNetworkingSecurityProduction TroubleshootingCoding for Automation
Soft Skills
Influencing Technical DirectionCross-Team Collaboration
Tools & Technologies
ECS/FargatePostgreSQLTypeScriptNode.jsOpenTelemetry
Certifications & Qualifications
ISO 27001Cyber Essentials Plus
Industry Keywords
Healthcare TechnologyDisaster RecoveryOperational MaturitySLO DefinitionIncident Detection

Tech Stack

Tools & technologies
AWSJavaScriptNode.jsPostgresTerraformTypeScript

About the role

Key responsibilities & impact
  • Own the reliability, security, and scalability of Lyrebird's production platform
  • Manage AWS infrastructure, deployment pipelines, observability, and incident detection and response
  • Define SLOs, alerting, and observability strategy
  • Hold the line on platform availability and latency for clinicians
  • Shape incident detection, response, and learning processes
  • Evolve AWS and Terraform infrastructure for resilience, scale, and disaster recovery
  • Set direction for CI/CD pipelines and developer tooling to make deployments safer and faster
  • Build security controls and compliance automation into the platform
  • Raise operational maturity across engineering teams so they can run their own services confidently
  • Challenge technical decisions when reliability or security is at stake
  • Report to the Engineering Manager, Platform and Security, and work across every engineering team

Requirements

What you’ll need
  • Deep hands-on experience running production workloads on AWS with Terraform, containers, and CI/CD
  • Strong grounding in networking, security, and production troubleshooting
  • Track record of designing observability, defining SLOs, and leading incident response that measurably improved reliability
  • Coding ability to automate operational work
  • Evidence of influencing technical direction across multiple engineering teams without relying on authority
  • Nice to have: Experience with ECS/Fargate, PostgreSQL, and distributed or asynchronous workloads
  • Nice to have: Familiarity with TypeScript/Node.js and OpenTelemetry
  • Nice to have: Exposure to healthcare technology or frameworks such as ISO 27001 or Cyber Essentials Plus
  • Nice to have: Experience operating platforms across multiple regions
  • Right to work in Australia on an ongoing basis without requiring visa sponsorship now or in the future

Benefits

Comp & perks
  • Inclusive, safe, supportive, and thriving workplace culture
  • Encouragement for applicants from underrepresented backgrounds in tech