FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and implementing scalable, reliable, and auditable systems, with a strong focus on operational excellence and quality engineering. Proficient in leading engineering teams while actively contributing to production code and testing frameworks.
Highest-signal resume keywords
Full Stack Software EngineeringPlatform Reliability EngineeringAutomated Testing FrameworksCloud-Native Applications on KubernetesJava, Python, Node.js, or TypeScript
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Full Stack Software EngineeringAutomated Testing FrameworksJavaPythonNode.jsTypeScriptReliability Patterns for Distributed SystemsStructured LoggingMetrics CollectionMonitoring and Alerting
Soft Skills
LeadershipMentoringCommunication
Tools & Technologies
KubernetesAI-Assisted Development ToolsGitHub Copilot
Certifications & Qualifications
Bachelor's Degree in Computer Science or Related Field
Industry Keywords
Operational ExcellenceQuality EngineeringProduction OperationsFault ToleranceResilience
Tech Stack
Tools & technologiesCloudDistributed SystemsJavaJavaScriptKubernetesNode.jsPythonTypeScript
About the role
Key responsibilities & impact- Architect and deliver foundational services for dependable, auditable, and scalable real-time recommendations
- Design and implement the State Machine for authoritative state and legal transitions
- Design and implement the Transactional Outbox for exactly-once intent emission to downstream consumers
- Own delivery for the platform scale and reliability pod, including planning, execution, quality, and operational readiness
- Own uptime, performance, resilience, and operational excellence across platform services
- Identify bottlenecks, failure points, and scaling constraints before production incidents
- Build and maintain automated unit, integration, end-to-end, regression, and performance testing frameworks
- Establish quality standards for delivery pods before release
- Write production code for reliability tooling, test automation, observability, and platform infrastructure
- Own logging, metrics, distributed tracing, alerting, and monitoring strategy
- Define patterns for exception handling, fault tolerance, recovery, and operational diagnostics
- Evaluate platform behavior under load and improve performance, throughput, reliability, and cost efficiency
- Define release readiness criteria, operational quality gates, and production handoff standards
- Manage and mentor engineers and contractors
- Conduct code reviews and uphold engineering standards
- Partner with decisioning, orchestration, activation, platform, and data teams
- Ensure services are production-ready, observable, resilient, and scalable across a distributed ecosystem
Requirements
What you’ll need- Bachelor's degree in computer science or related field
- 8+ years of full stack software engineering experience
- At least 1–2 years leading platform reliability, quality engineering, site reliability, or production operations initiatives
- Strong experience with Java, Python, Node.js, or TypeScript
- Experience designing and implementing automated testing frameworks, including integration, end-to-end, load, and quality automation testing
- Strong understanding of structured logging, metrics collection, distributed tracing, monitoring, and alerting
- Experience designing reliability patterns for distributed systems, including fault tolerance, retries, resilience, and failure recovery
- Ability to lead a small engineering pod and manage contractor resources while remaining an active individual contributor
- Experience operating cloud-native applications on Kubernetes and container-based platforms
- Ability to articulate operational risks, quality concerns, and engineering tradeoffs to technical leaders and senior executives
- Experience using AI-assisted development tools such as Claude or GitHub Copilot
- Qualified candidates must currently live in, or be willing to move to, a commutable distance for a hybrid work arrangement
- Ability to work Monday-Friday, 8 hours per day, 5 days per week
- Home internet with at least 25 Mbps download and 10 Mbps upload speed
- Dedicated workspace without ongoing interruptions to protect member PHI/HIPAA information
Benefits
Comp & perks- Bonus incentive plan based on company and/or individual performance
- Medical benefits
- Dental benefits
- Vision benefits
- 401(k) retirement savings plan
- Paid time off
- Company holidays
- Personal holidays
- Paid parental leave
- Paid caregiver leave
- Short-term disability
- Long-term disability
- Life insurance
- Home or hybrid home/office work arrangement
- Occasional travel for training or meetings
- Flexible business hours may be possible depending on business needs
