Lead Site Reliability Engineer

JPMorgan Chase

Confirmed live yesterday Low trust

Quick summary

Work type
On-site
Location
Jersey City, NJ
Posted
43 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $181k
$133k most similar roles pay here $235k

This listing doesn't post a salary. Most similar roles pay $147,480–$213,692.

Based on 238 similar postings.

Employer

About JPMorgan Chase

JPMorgan Chase & Co. is a global financial services firm and one of the largest banks in the world, offering investment banking, commercial banking, asset management, and consumer financial services.

JPMorgan Chase currently has 1117 open roles on FindRole.

Listed pay typically runs $186,160–$215,000 across 7 roles with salary data.

Most-posted roles

View all roles at JPMorgan Chase

At a glance

TL;DR · Lead Site Reliability Engineer

As a Lead Site Reliability Engineer within the Corporate Technology /Infrastructure Platforms Client Solutions team, you will hold a leadership role providing technical expertise and mentorship to other engineers. You will lead resiliency design reviews, break down complex problems into manageable tasks, and champion site reliability culture through data-driven analytics to improve service levels. Your daily work involves identifying bottlenecks, establishing service level objectives, and managing major incidents to prevent financial loss. You will integrate enterprise-authorized AI capabilities into the software development lifecycle for incident triage, troubleshooting, and automated quality checks while ensuring security compliance. Required skills include proficiency in Python, Java/Spring Boot, or .Net, along with experience in container orchestration, CI/CD pipelines, observability, and networking. The role focuses on maintaining high availability and stability for large-scale financial services technology platforms.

What you'll do

  • Lead resiliency design reviews and break down complex problems into manageable tasks for other engineers.
  • Use data-driven analytics to identify and resolve technology bottlenecks to improve platform stability.
  • Define service level indicators, objectives, and error budgets in collaboration with stakeholder partners.
  • Act as the primary point of contact during major incidents to quickly resolve issues and prevent financial loss.
  • Provide technical expertise, mentorship, and guidance to other engineers across multiple technical domains.
  • Integrate AI-assisted workflows into SDLC practices including CI/CD quality checks and test automation.
  • Evaluate AI-assisted operational recommendations for accuracy while ensuring compliance with security and risk guardrails.

What we're looking for

  • Formal training or certification on site reliability engineering concepts and 5+ years of applied experience.
  • Demonstrated proficiency in reliability, scalability, performance, security, enterprise system architecture, and toil reduction.
  • Fluency in at least one programming language such as Python, Java/Spring Boot, or .Net.
  • Experience using enterprise-authorized AI capabilities to improve SRE workflows while ensuring data sensitivity and security.
  • Ability to evaluate AI-assisted operational recommendations for correctness and define guardrails for team usage.
  • Proficiency in observability, including white and black box monitoring, SLO alerting, and telemetry collection.
  • Proficiency with containerization, orchestration, and CI/CD practices and tooling.
  • Experience troubleshooting common networking technologies and issues.

More like this

Similar roles

Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 44 days ago
Site Reliability Engineering Python Java Spring Boot .NET CI/CD Docker Kubernetes Terraform AWS Prometheus Grafana Dynatrace Datadog Splunk infrastructure-as-code CloudFormation ECS Incident Management Observability
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 45 days ago
Site Reliability Engineering Python Java Spring Boot .Net AI CI/CD Container Orchestration AWS Observability Monitoring Telemetry Networking System Architecture SDLC
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

New York, NY 43 days ago
Site Reliability Engineering SRE CI/CD Observability Monitoring Automation Change Management Dynatrace Splunk Geneos Grafana ITIL AWS Azure GCP Python Shell PowerShell Ansible Terraform Kubernetes OpenShift Microservices
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 43 days ago
SRE CI/CD Jenkins GitLab Terraform Docker Kubernetes ECS AI Python Go JavaScript GraphQL Kafka OpenTelemetry Networking
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 45 days ago
Site Reliability Engineering Observability Monitoring Telemetry Service Level Objectives Alerting AI SDLC Automation
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Palo Alto, CA 59 days ago
Site Reliability Engineering Java Go Python Terraform Kubernetes Docker CI/CD GitOps Grafana Prometheus Dynatrace Datadog Splunk Kafka RabbitMQ SQS Neo4j Pinecone Weaviate Chroma LangChain LangGraph AutoGen CrewAI GitHub Copilot Fluentd Logstash Vector RESTful APIs RAG TensorFlow PyTorch scikit-learn Hadoop Spark Flink MongoDB Cassandra DynamoDB InfluxDB TimescaleDB AWS Azure GCP
5+ yrs exp