Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Databento

Site Reliability Engineer

Databento

. Own uptime, SLAs, and SLOs across API and platform services .

Posted 10/1/2026full-timeRemote • United StatesMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in Site Reliability Engineering (SRE) and DevOps practices, with a strong focus on Python application optimization, observability tooling, and high-availability deployment strategies. Proven ability to manage incident response and improve CI/CD workflows in a remote work environment.

Highest-signal resume keywords
Site Reliability Engineering (SRE)Python Application DevelopmentObservability ToolingContainerization and High-Availability DeploymentIncident Response Management

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
PythonDebugging ToolsPerformance OptimizationDatabase Schema DesignQuery Optimization
Soft Skills
Good Communication SkillsStrong Work Ethic
Tools & Technologies
PrometheusOpenTelemetryDockerKubernetesAnsible
Industry Keywords
API ServicesOperational Best PracticesPetabyte-Scale Data ProcessingFinancial DataAlgorithmic Trading

Tech Stack

Tools & technologies
AnsibleDockerKubernetesLinuxLogstashPrometheusPythonTerraform

About the role

Key responsibilities & impact
  • Own uptime, SLAs, and SLOs across API and platform services
  • Set reliability and operational best practices for developers
  • Build and maintain observability across logging, metrics, and tracing
  • Design and run high-availability deployment and containerization strategies
  • Profile and optimize Python applications for throughput, latency, and cost
  • Debug production issues down to the OS level using strace, perf, eBPF, ss, and gdb
  • Improve deployment and CI/CD workflows
  • Participate in the on-call rotation, lead incident response, and run post-incident reviews
  • Identify needed fixes and take projects from idea to completion
  • Work on petabyte-scale data processing, customer management and billing, and query systems powering APIs

Requirements

What you’ll need
  • Midlevel or senior individual contributor
  • Full-time experience in SRE, DevOps, or backend engineering, preferably at a trading firm, tech company, or high-growth startup
  • Hands-on experience with observability tooling for logging, metrics, and tracing, such as Prometheus, OpenTelemetry, VictoriaMetrics, Jaeger, Logstash, Loki, or Vector
  • Experience with containerization and high-availability deployment, such as Docker, Podman, Docker Compose, Docker Swarm, Kubernetes, or k3s
  • Strong proficiency in Python, including application development and performance optimization
  • Comfortable with Linux debugging and profiling tools such as strace, perf, eBPF, ss, and gdb
  • Track record of measurable impact in a recent role
  • Experience with alerting and incident response best practices is a plus
  • Familiarity with configuration management or infrastructure-as-code tools such as Ansible or Terraform is helpful
  • HTTP benchmarking, load testing, and capacity planning experience is a bonus
  • Database schema design and query optimization skills are nice to have
  • Good communication skills and work ethic for a remote workplace
  • Interest in financial data or algorithmic trading

Benefits

Comp & perks
  • Equal employment opportunities and nondiscrimination protections
  • Employment accommodation available upon request
  • AI-powered Talent Matching opt-out option