FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing and supporting cloud-based production environments, with a strong focus on uptime, performance tuning, and automation. Proficient in utilizing cutting-edge technologies and fostering a collaborative culture of site reliability and accountability.
Highest-signal resume keywords
Site Reliability EngineeringCloud-Based Infrastructure ManagementAWS/GCP ExperienceProduction Environment SupportDevOps Culture Enforcement
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Distributed Production Systems DevelopmentPerformance TuningAutomation and Tooling DevelopmentDebugging Production IssuesNetworking Protocols (TCP/IP, HTTP/HTTPS, TLS, DNS, NTP)
Soft Skills
Highly CollaborativeAutonomousIndividually AccountableCommitment to Diversity and Inclusion
Tools & Technologies
DockerKubernetesLoad BalancersReverse ProxiesWeb Application Firewalls
Industry Keywords
Consumer Facing Web Applications24/7 Cloud-Based Production EnvironmentBlameless Post-Mortems
Tech Stack
Tools & technologiesAWSCloudDNSDockerFirewallsGoogle Cloud PlatformKubernetesTCP/IP
About the role
Key responsibilities & impact- Develop, monitor, and maintain distributed production systems
- Build frameworks and processes for ensuring uptime for patients and providers
- Monitor and maintain complex cloud-based infrastructure, systems, and services
- Automate and develop tooling, processes, and infrastructure to make development faster, repeatable, and error-proof
- Support product engineering teams with scaling, performance, and uptime needs
- Diagnose and debug production-related issues
- Analyze and performance-tune systems, code, and networking for scaling and optimal operation
- Work with cutting-edge GenAI tools and technology
- Enforce a culture of strong DevOps and shared product-team responsibility for site reliability and first response
Requirements
What you’ll need- 5+ years of supporting consumer facing web application production environments and systems in a Site Reliability Engineering or Production Engineering role
- 2+ years of on-call experience in a 24/7 cloud-based production environment
- 2+ years of experience in managing and supporting modern cloud-based environments and infrastructure like AWS/GCP, Docker, Kubernetes, etc.
- Experience with edge technologies such as load balancers, reverse proxies, web application firewalls, routing, etc.
- Deep understanding of protocols such as TCP/IP, HTTP/HTTPS, TLS, DNS, NTP
- Bachelor’s degree in Computer Science, Computer Engineering, or equivalent engineering experience is a plus, but not required
- Comfortable in an outage situation and believe in blameless post-mortems
- Highly collaborative, autonomous, individually accountable, and committed to diverse and inclusive teams
Benefits
Comp & perks- Flexible, hybrid work environment at our convenient Soho location (If based in NYC)
- Unlimited Vacation
- 100% paid employee health benefit options (including medical, dental, and vision)
- Commuter Benefits
- 401(k) with employer funded match
- Corporate wellness program with Wellhub
- Sabbatical leave (for employees with 5+ years of service)
- Competitive paid parental leave and fertility/family planning reimbursement
- Cell phone reimbursement
- Catered lunch everyday along with beverages and snacks
- Employee Resource Groups and ZocClubs to promote shared community and belonging
- Great Place to Work Certified
