Lead Site Reliability Engineer

JPMorgan Chase

Confirmed live today High trust

Quick summary

Work type
On-site
Location
Houston, TX
Employment
Full-time
Posted
today
Freshness
Confirmed live today

Market check

Salary context

How this pay compares to similar roles

Similar $182k
$136k most similar roles pay here $237k

This listing doesn't post a salary. Most similar roles pay $146,837–$217,500.

Based on 240 similar postings.

Employer

About JPMorgan Chase

JPMorgan Chase & Co. is a global financial services firm and one of the largest banks in the world, offering investment banking, commercial banking, asset management, and consumer financial services.

JPMorgan Chase currently has 1164 open roles on FindRole.

Most-posted roles

View all roles at JPMorgan Chase

At a glance

TL;DR · Lead Site Reliability Engineer

As a Lead Site Reliability Engineer within the Corporate & Investment Bank Management and Support Functions Digital & Platform Services team, you will champion site reliability culture while providing high-level technical expertise and mentorship. You will lead initiatives to improve application stability using data-driven analytics, identify bottlenecks, and collaborate with stakeholders to establish service level objectives and error budgets. Your daily work involves managing major incidents, implementing reuse-first AI-assisted reliability workflows across the SDLC, and ensuring security controls in CI/CD pipelines. The role requires proficiency in Python, Java/Spring Boot, or .Net, along with expertise in observability, telemetry collection, container orchestration, and networking technologies. You will solve complex problems regarding scalability, performance, and toil reduction while utilizing enterprise-authorized AI tools to accelerate incident triage and automate operational readiness within a high-stakes financial services environment.

What you'll do

  • Champion site reliability culture and share knowledge through internal forums and communities of practice.
  • Use data-driven analytics to identify and resolve technology bottlenecks to improve service levels.
  • Establish service level indicators, objectives, and error budgets with stakeholder partners.
  • Utilize enterprise AI tools to accelerate major-incident triage, troubleshooting, and post-incident analysis.
  • Serve as the primary point of contact during major incidents to resolve issues quickly and prevent financial loss.
  • Provide technical expertise and mentorship to other engineers across various technical domains.
  • Implement AI-assisted reliability workflows across SDLC practices including CI/CD quality checks and test automation.

What we're looking for

  • Formal training or certification on site reliability engineering concepts.
  • 5+ years of applied experience in site reliability engineering.
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, or .Net.
  • Experience using enterprise-authorized AI capabilities to improve SRE workflows and triage incidents.
  • Ability to evaluate AI-assisted operational recommendations for correctness, risk, and security.
  • Proficiency in observability, including white and black box monitoring, SLO alerting, and telemetry collection.
  • Proficiency with continuous integration and continuous delivery practices and tooling.
  • Experience with container orchestration, networking troubleshooting, and advanced software application knowledge (preferred).

More like this

Similar roles

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 70 days ago
Site Reliability Engineering Python Java Spring Boot .Net Kubernetes AWS Google Cloud CI/CD Observability Monitoring Telemetry Networking Infrastructure Optimization FinOps Disaster Recovery Capacity Management
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 72 days ago
Site Reliability Engineering Python Java Spring Boot .Net AI CI/CD Container Orchestration AWS Observability Monitoring Telemetry Networking System Architecture SDLC
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

OH 19 days ago
Site Reliability Engineering Python Go Java C++ Rust Kubernetes Terraform CI/CD Prometheus Grafana Splunk Datadog Dynatrace Networking AI Prompt Engineering Agent Orchestration Infrastructure Automation
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

New York, NY 70 days ago
Site Reliability Engineering SRE CI/CD Observability Monitoring Automation Change Management Dynatrace Splunk Geneos Grafana ITIL AWS Azure GCP Python Shell PowerShell Ansible Terraform Kubernetes OpenShift Microservices
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 22 days ago
SRE AI CI/CD Dynatrace Splunk Geneos Grafana ITIL AWS Azure GCP Python Shell PowerShell Ansible Terraform Kubernetes OpenShift Microservices
10+ yrs exp

Lead Site Reliability Engineer

Mastercard

O Fallon, MO 23 days ago $122,000–$207,000
Site Reliability Engineering Java Spring Framework Python Go DevOps CI/CD Configuration Management Distributed Systems Automation Observability ITSM Root Cause Analysis Capacity Planning Monitoring