Lead Site Reliability Engineer

JPMorgan Chase

Confirmed live 2 days ago Low trust

Quick summary

Work type
On-site
Location
Plano, TX
Posted
43 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

How this pay compares to similar roles

Similar $181k
$133k most similar roles pay here $235k

This listing doesn't post a salary. Most similar roles pay $147,480–$213,692.

Based on 238 similar postings.

Employer

About JPMorgan Chase

JPMorgan Chase & Co. is a global financial services firm and one of the largest banks in the world, offering investment banking, commercial banking, asset management, and consumer financial services.

JPMorgan Chase currently has 1117 open roles on FindRole.

Listed pay typically runs $186,160–$215,000 across 7 roles with salary data.

Most-posted roles

View all roles at JPMorgan Chase

At a glance

TL;DR · Lead Site Reliability Engineer

As a Lead Site Reliability Engineer within the Identity & Access Management team, you will serve as the non-functional requirement owner for critical applications. You will lead SRE practices to balance delivery speed with system stability while driving improvements in resiliency, security, scalability, monitoring, and automation. Your daily responsibilities include partnering with stakeholders to scale SRE adoption, conducting blameless post-incident reviews, and coaching junior engineers. You will integrate enterprise-authorized AI capabilities into the software development lifecycle for incident triage and automated quality checks. The role requires expertise in CI/CD tools like Jenkins, GitLab, and Terraform, alongside container orchestration using Docker, Kubernetes, or ECS. Technical requirements include proficiency in JavaScript, Go, or Python, knowledge of GraphQL, event-driven architecture via Kafka, OpenTelemetry, and a deep understanding of distributed systems and networking technologies.

What you'll do

  • Lead SRE practices to balance delivery speed, efficiency, and system stability.
  • Drive improvements in scalability, monitoring, instrumentation, and automation for assigned applications.
  • Establish reliability expectations and track progress through specific stability metrics.
  • Conduct blameless, data-driven post-incident reviews to turn lessons into actionable improvements.
  • Scale SRE adoption across various application and platform teams.
  • Integrate AI-assisted workflows into the SDLC and toolchain for automated testing and operational readiness.
  • Use enterprise-authorized AI tools to accelerate incident triage and troubleshooting while ensuring security compliance.
  • Coach and mentor entry-to-mid-level engineers through knowledge sharing and internal forums.

What we're looking for

  • Formal training or certification in software engineering concepts plus 5+ years of applied experience.
  • Advanced knowledge of SRE principles and a track record of implementing SRE across application and platform teams.
  • Experience leading technologists to manage and resolve complex technology issues at a firmwide level.
  • Experience hiring, developing, and recognizing talent.
  • Hands-on experience with CI/CD tools such as Jenkins, GitLab, or Terraform.
  • Experience with containers and orchestration including Docker, Kubernetes, or ECS.
  • Demonstrated experience using enterprise-authorized AI capabilities to improve SRE workflows while ensuring data security.
  • Ability to evaluate AI-assisted operational recommendations for correctness and risk (preferred); proficiency in JavaScript, Go, or Python (preferred).

More like this

Similar roles

Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 44 days ago
Site Reliability Engineering Python Java Spring Boot .NET CI/CD Docker Kubernetes Terraform AWS Prometheus Grafana Dynatrace Datadog Splunk infrastructure-as-code CloudFormation ECS Incident Management Observability
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 43 days ago
Site Reliability Engineering Python Java Spring Boot .Net Kubernetes AWS Google Cloud CI/CD Observability Monitoring Telemetry Networking Infrastructure Optimization FinOps Disaster Recovery Capacity Management
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 45 days ago
Site Reliability Engineering Python Java Spring Boot .Net AI CI/CD Container Orchestration AWS Observability Monitoring Telemetry Networking System Architecture SDLC
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

New York, NY 43 days ago
Site Reliability Engineering SRE CI/CD Observability Monitoring Automation Change Management Dynatrace Splunk Geneos Grafana ITIL AWS Azure GCP Python Shell PowerShell Ansible Terraform Kubernetes OpenShift Microservices
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 45 days ago
Site Reliability Engineering Observability Monitoring Telemetry Service Level Objectives Alerting AI SDLC Automation
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Palo Alto, CA 59 days ago
Site Reliability Engineering Java Go Python Terraform Kubernetes Docker CI/CD GitOps Grafana Prometheus Dynatrace Datadog Splunk Kafka RabbitMQ SQS Neo4j Pinecone Weaviate Chroma LangChain LangGraph AutoGen CrewAI GitHub Copilot Fluentd Logstash Vector RESTful APIs RAG TensorFlow PyTorch scikit-learn Hadoop Spark Flink MongoDB Cassandra DynamoDB InfluxDB TimescaleDB AWS Azure GCP
5+ yrs exp