Reliability Engineer IV, Observability Specialist

US Bank

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Chicago, ILAtlanta, GACupertino, CAGresham, ORDenver, COCharlotte, NCBrookfield, WIIrving, TXHopkins, MNEarth City, MO
Salary
$124,355–$146,300 / yr
Employment
Full-time
Posted
3 days ago
Freshness
Confirmed live yesterday
Closes
Oct 12, 2026

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $188k
This role $135k
$112k most similar roles pay here $237k

This role pays less than 87% of similar roles. Most pay $163,000–$214,000 — the shaded band above. At the midpoint, this role pays about $135k versus about $188k for comparable roles.

Based on 240 similar postings.

Employer

About US Bank

U.S. Bank (U.S. Bancorp) is the fifth-largest bank in the United States, providing retail banking, corporate and commercial banking, wealth management, and payment services to millions of customers. Industry: Banking & Financial Services

US Bank currently has 28 open roles on FindRole.

Listed pay typically runs $115,685–$136,100 across 22 roles with salary data.

Most-posted roles

View all roles at US Bank

At a glance

TL;DR · Reliability Engineer IV, Observability Specialist

Reliability Engineer 4 (Observability Specialist) serves as a senior-level technical leader focused on translating customer journeys into measurable reliability objectives. Working with product owners, application engineering, and SRE teams, the role establishes governance for SLIs, SLOs, error budgets, and telemetry standards to ensure service reliability. Day-to-day responsibilities include designing observability architectures, managing synthetic monitoring, developing executive health dashboards, and performing incident analysis to reduce alert fatigue. The position requires expertise in distributed systems, microservices, and cloud platforms like Kubernetes. Candidates should possess hands-on experience with tools such as Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry. This role addresses the critical business problem of ensuring production-ready applications remain reliable by providing high-quality data for detection and triage while aligning technical monitoring strategies with specific customer outcomes and operational risk management.

What you'll do

  • Translate customer journeys and business outcomes into measurable reliability objectives like SLIs and SLOs.
  • Establish and maintain governance frameworks for dashboards, alerts, synthetic monitoring, and telemetry standards.
  • Design and implement scalable observability architectures including instrumentation standards and tagging frameworks.
  • Develop executive and operational service health dashboards to monitor availability, latency, and customer impact.
  • Analyze telemetry data and incident trends to identify gaps and reduce alert fatigue.
  • Provide technical leadership and mentorship on distributed tracing, logging, and performance monitoring.
  • Maintain an authoritative inventory of all observability assets and ensure compliance with established standards.

What we're looking for

  • Bachelor's degree or equivalent work experience.
  • Six to eight years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application development.
  • Expertise in Observability Engineering, Site Reliability Engineering (SRE), or Reliability Engineering (preferred).
  • Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring (preferred).
  • Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks (preferred).
  • Proficiency with Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry (preferred).
  • Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes (preferred).
  • Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements (preferred).

More like this

Similar roles

Senior Engineer, Reliability

LPL Financial

Austin, TX +1 21 days ago $101,558–$169,229
SRE Observability AWS Dynatrace ELK ServiceNow SolarWinds CI/CD Release Management Change Management Monitoring Automation SDLC Incident Management Runbooks
6+ yrs exp

Senior Data Engineer, Observability Engineering

CVS Health

Remote (AZ) 11 days ago $92,700–$222,480
Databricks PySpark Spark SQL Python Delta Lake Unity Catalog Structured Streaming CI/CD GitHub Actions Azure DevOps OCSF Telemetry Distributed Data Processing Data Governance Metadata Management
5+ yrs exp Remote

Executive Director, Site Reliability Engineering, Retail Pharmacy

CVS Health

Remote 10 days ago $175,100–$334,750
SRE Site Reliability Engineering AWS Microsoft Azure Google Cloud Platform Kubernetes OpenShift Splunk Dynatrace Datadog Prometheus Grafana AI Machine Learning Distributed Systems Edge Computing HIPAA PCI DSS Adobe Analytics
10+ yrs exp Remote

Reliability Engineer

Anduril Industries

Atlanta, GA +1 100 days ago $126,000–$167,000
FMEA FTA Weibull Analysis MIL-HDBK-217 MIL-HDBK-472 MIL-STD-810 MIL-STD-461 MIL-STD-516C MIL-STD-1629 HALT HASS HITL SITL Root Cause Analysis CAPA Technical Writing Data Acquisition Environmental Testing Mechanical Testing
5+ yrs exp

Reliability Engineer

Anduril Industries

Costa Mesa, CA +1 100 days ago $146,000–$194,000
FMEA FTA Weibull Analysis MIL-HDBK-217 MIL-HDBK-472 MIL-STD-810 MIL-STD-461 MIL-STD-516C MIL-STD-1629 HALT HASS HITL SITL Root Cause Analysis CAPA Data Acquisition Technical Writing Systems Engineering
5+ yrs exp