FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Tech Stack
Tools & technologiesCloudDistributed SystemsGrafanaITSMKubernetesMicroservicesPrometheusServiceNowSplunk
About the role
Key responsibilities & impact- Support production applications and proactively automate discoveries
- Eliminate recurring incidents and reduce time to restore customer service
- Improve application availability, latency, performance, efficiency, and proactive monitoring
- Interface with business users, development teams, and system administrators
- Develop, coordinate, and conduct technical reliability studies on engineering designs
- Measure and analyze reliability of designs, materials, processes, costs, and final production products
- Recommend design or test methods and statistical process control procedures to achieve required reliability levels
- Complete risk analysis studies of new designs and processes
- Test and analyze failures and propose design or formulation changes to improve system and process reliability
Requirements
What you’ll need- Bachelor's degree, or equivalent work experience
- Five to seven years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application development
- Expertise in Site Reliability Engineering (SRE) or Reliability Engineering
- Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring
- Ability to understand stakeholder needs and guide development of reliability requirements for large, complex multi-system products
- Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks
- Proficiency with Datadog, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry
- Experience building, standardizing, and tuning operational dashboards and actionable alerts
- Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes
- Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements
- Excellent stakeholder management, communication, and technical leadership skills
- Hands-on experience with ServiceNow
- Ability to work from a U.S. Bank location three or more days per week
- Ability to comply with U.S. Bank policies and procedures, including the Code of Ethics and Business Conduct and workplace conduct and safety policies
Benefits
Comp & perks- Healthcare (medical, dental, vision)
- Basic term and optional term life insurance
- Short-term and long-term disability
- Pregnancy disability and parental leave
- 401(k) and employer-funded retirement plan
- Paid vacation (from two to five weeks depending on salary grade and tenure)
- Up to 11 paid holiday opportunities
- Adoption assistance
- Sick and Safe Leave accruals of one hour for every 30 worked, up to 80 hours per calendar year unless otherwise provided by law
- Incentive and recognition programs
- Equity stock purchase
- 401(k) contribution
- Pension
