Find Jobs
Find Jobs Near You – Available Work in Your Location
Skip to job details
UB
U.S. Bank National Association
Reliability Observability Engineer 2
Choose a Location
This role is available in multiple locations. Pick one to apply.
Career Insights for Systems Engineer
See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.
Scorecard
Based on Colorado data
Review key factors to help you decide if this role fits your goals. How is this calculated?
What they do
A Systems Engineer creates computer and data communication networks for companies and organizations. Plans and designs layout for a network, determines the hardware needed and placement of computers, servers, cables and routers; determines data storage, system capacity, and speed.
$113,680 / year median in Colorado
+7% projected growth
Job Description
Reliability Observability Engineer 2 U.S. Bank National Association - 3.5 Englewood, CO Job Details $86,360 - $101,600 a year 1 day ago Benefits Paid holidays Disability insurance Health insurance Dental insurance 401(k) Adoption assistance Parental leave Vision insurance Life insurance Qualifications Risk analysis Stakeholder relationship building Stakeholder management Full Job Description At U.S. Bank, we're on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions, enabling the communities we support to grow and succeed in the right ways, all more confidently and more often—that's what we call the courage to thrive. We believe it takes all of us to bring our shared ambition to life, and each person is unique in their potential. A career with U.S. Bank gives you a wide, ever-growing range of opportunities to discover what makes you thrive. Try new things, learn new skills and discover what you excel at—all from Day One. As a wholly owned subsidiary of U.S. Bank, Elavon is committed to building the platforms and ecosystems that help over 1.5 million customers around the world to achieve their financial goals—no matter what they need. From transaction processing to customer service, to driving innovation and launching new products, we're building a range of tailored payment solutions powered by the latest technology. As part of our team, you can explore what motivates and energizes your career goals: partnering with our customers, our communities, and each other. Job Description Job Description Key Responsibilities Lead Observability Strategy across critical customer journeys, aligning monitoring capabilities with business outcomes, reliability goals, and customer experience. Define, implement, and govern SLIs, SLOs, Error Budgets, and Reliability Metrics for enterprise applications and services. Design and maintain scalable Observability Architectures , including telemetry instrumentation, monitoring frameworks, tagging standards, and alerting models. Establish Observability Governance for dashboards, alerts, synthetic monitoring, telemetry standards, and lifecycle management of monitoring assets. Partner with Product, Engineering, SRE, and Operations teams to ensure production readiness, application instrumentation, and reliability measurement. Develop and optimize Service Health Dashboards and Reporting that provide visibility into availability, latency, customer impact, dependency performance, and SLO compliance. Analyze telemetry data, incidents, alert history, problem records, and performance trends to identify gaps, reduce alert fatigue, and improve detection accuracy. Provide technical leadership and mentorship on Distributed Tracing, Logging, Metrics, Synthetic Monitoring, Application Performance Monitoring (APM), Real User Monitoring (RUM), and Alert Governance best practices. Basic Qualifications Bachelor's degree, or equivalent work experience Four to five years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application development Preferred Skills/Experience Expertise in Observability Engineering , Site Reliability Engineering (SRE) , or Reliability Engineering . Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring . Demonstrated ability to understand stakeholder needs and guide the development of reliability requirements for large, complex multi-system products. Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks . Proficiency with Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry . Experience building, standardizing, and tuning operational dashboards and actionable alerts that communicate service health, customer impact, dependency health, performance trends, failure conditions, severity, ownership, routing, and runbook linkage. Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes . Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements. Excellent stakeholder management, communication, and technical leadership skills. Location expectations This role requires working from a U.S. Bank location three (3) or more days per week. If there's anything we can do to accommodate a disability during any portion of the application or hiring process, please refer to our disability accommodations for applicants.