Reliability Engineer 3, Observability Specialist

US Bank

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
Brookfield, WIAtlanta, GAHopkins, MNCupertino, CACharlotte, NC
Salary
$98,175–$115,500 / yr
Posted
8 days ago
Freshness
Confirmed live 2 days ago
Closes
Sep 28, 2026

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $172k
This role $107k
$84k most similar roles pay here $229k

This role pays less than 95% of similar roles. Most pay $136,493–$208,350 — the shaded band above. At the midpoint, this role pays about $107k versus about $172k for comparable roles.

Based on 240 similar postings.

Employer

About US Bank

U.S. Bank (U.S. Bancorp) is the fifth-largest bank in the United States, providing retail banking, corporate and commercial banking, wealth management, and payment services to millions of customers. Industry: Banking & Financial Services

US Bank currently has 30 open roles on FindRole.

Listed pay typically runs $119,765–$140,900 across 29 roles with salary data.

Most-posted roles

View all roles at US Bank

At a glance

TL;DR · Reliability Engineer 3, Observability Specialist

Reliability Engineer 3 (Observability Specialist) joins the team to own and drive the enterprise Observability Strategy across multiple platforms and technology domains. The role involves defining and governing SLIs, SLOs, Error Budgets, and reliability standards while architecting scalable solutions for telemetry, distributed tracing, logging, metrics, synthetic monitoring, and APM/RUM capabilities. You will establish governance frameworks, manage dashboard lifecycles, and serve as a technical advisor to leadership on production readiness and risk management. The position requires expertise in Site Reliability Engineering and proficiency with tools such as Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry. You will solve complex problems regarding service health, latency, and availability within distributed systems, microservices, and cloud platforms to improve overall system resilience and reduce operational noise through advanced analysis of incident data.

What you'll do

  • Own and drive the enterprise Observability Strategy across multiple platforms and technology domains.
  • Define and govern enterprise-wide SLIs, SLOs, Error Budgets, and reliability standards.
  • Architect scalable observability solutions using telemetry, distributed tracing, logging, metrics, and synthetic monitoring.
  • Establish and oversee governance frameworks for instrumentation standards, telemetry policies, and alert management.
  • Advise leadership on technology strategy, production readiness, and operational risk management.
  • Design and evolve executive reporting for service health, availability, latency, and reliability trends.
  • Analyze incident data and telemetry to reduce operational noise and improve service resiliency.
  • Provide technical leadership and mentorship for observability engineering practices across the enterprise.

What we're looking for

  • Bachelor's degree or equivalent work experience.
  • Five to seven years of relevant experience in business/risk analysis, IT Service Management, production support, project management, or application development.
  • Expertise in Observability Engineering, Site Reliability Engineering (SRE), or Reliability Engineering.
  • Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring.
  • Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks.
  • Proficiency with tools such as Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry.
  • Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes.
  • Excellent stakeholder management, communication, and technical leadership skills.

More like this

Similar roles

Reliability Engineer III, Observability Specialist

US Bank

Brookfield, WI +4 16 days ago $98,175$115,500
Observability Site Reliability Engineering (SRE) SLIs SLOs Error Budgets Distributed Tracing Prometheus Grafana Datadog Dynatrace Splunk New Relic Elastic OpenTelemetry Kubernetes Microservices APM RUM Synthetic Monitoring
5+ yrs exp

Senior Observability Engineer

RBC

Minneapolis, MN 29 days ago $90,000$140,000
ELK Stack Dynatrace Prometheus Grafana Open Telemetry Jaeger Python Java Go Linux Unix Kubernetes OpenShift Tableau Anthropic OpenAI Copilot SRE

Site Reliability Engineering Lead

US Bank

Atlanta, GA +2 3 days ago $111,605$131,300
SRE DevOps AWS Azure Kubernetes Docker Terraform Ansible Python PowerShell Shell Scripting CI/CD GitHub Actions Azure DevOps Jenkins GitLab Datadog Splunk Dynatrace Grafana Prometheus CloudWatch Azure Monitor OpenTelemetry ServiceNow Jira REST APIs SQL
6+ yrs exp

Reliability Engineer

Anduril Industries

Atlanta, GA +1 79 days ago $126,000$167,000
FMEA FTA Weibull Analysis MIL-HDBK-217 MIL-HDBK-472 MIL-STD-810 MIL-STD-461 MIL-STD-516C MIL-STD-1629 HALT HASS HITL SITL Root Cause Analysis CAPA Technical Writing Data Acquisition Environmental Testing Mechanical Testing
5+ yrs exp

Reliability Engineer

Anduril Industries

Costa Mesa, CA +1 79 days ago $146,000$194,000
FMEA FTA Weibull Analysis MIL-HDBK-217 MIL-HDBK-472 MIL-STD-810 MIL-STD-461 MIL-STD-516C MIL-STD-1629 HALT HASS HITL SITL Root Cause Analysis CAPA Data Acquisition Technical Writing Systems Engineering
5+ yrs exp