About The Position

At U.S. Bank, we’re on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions and enabling the communities we support to grow and succeed. We believe it takes all of us to bring our shared ambition to life, and each person is unique in their potential. A career with U.S. Bank gives you a wide, ever-growing range of opportunities to discover what makes you thrive at every stage of your career. Try new things, learn new skills and discover what you excel at—all from Day One. As a Reliability Engineer, your role will be a combination of supporting production applications and proactively looking for ways to automate your discoveries, eliminate incidents from recurring and/or reduce the time it takes to get our customers back up and running. In addition, you'll focus on improving the following for our applications: availability, latency, performance, efficiency, and effective proactive monitoring. The reliability engineer interfaces with business users, development teams and system administrators to ensure systems perform to meet their business needs and specifications.

Requirements

  • Bachelor's degree, or equivalent work experience
  • Five to seven years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application development

Nice To Haves

  • Expertise in Site Reliability Engineering (SRE), or Reliability Engineering.
  • Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring.
  • Demonstrated ability to understand stakeholder needs and guide the development of reliability requirements for large, complex multi-system products.
  • Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks.
  • Proficiency with Datadog, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry.
  • Experience building, standardizing, and tuning operational dashboards and actionable alerts that communicate service health, customer impact, dependency health, performance trends, failure conditions, severity, ownership, routing, and runbook linkage.
  • Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes.
  • Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements.
  • Excellent stakeholder management, communication, and technical leadership skills.
  • Hands on experience with ServiceNow.

Responsibilities

  • Developing, coordinating, and conducting technical reliability studies on engineering designs to assess the likelihood that a product/process performs its intended function over the intended lifecycle.
  • Measuring and analyzing the reliability of the design, materials, processes, cost, and final products of production.
  • Recommending design or test methods and statistical process control procedures for achieving required levels of product reliability.
  • Completing risk analysis studies of new designs and processes.
  • Undertaking testing and analysis on failures, proposing changes in design or formulation to improve system and/or process reliability.

Benefits

  • Healthcare (medical, dental, vision)
  • Basic term and optional term life insurance
  • Short-term and long-term disability
  • Pregnancy disability and parental leave
  • 401(k) and employer-funded retirement plan
  • Paid vacation (from two to five weeks depending on salary grade and tenure)
  • Up to 11 paid holiday opportunities
  • Adoption assistance
  • Sick and Safe Leave accruals of one hour for every 30 worked, up to 80 hours per calendar year unless otherwise provided by law
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service