Site Reliability Engineer, Global Banking & Markets, Vice President

Goldman Sachs

Confirmed live 2 days ago Trusted

Quick summary

Work type
On-site
Location
New York, NY
Salary
$150,000–$250,000 / yr
Posted
37 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $188k
This role $200k
$136k most similar roles pay here $262k

This role pays more than 63% of similar roles. Most pay $159,323–$217,556 — the shaded band above. At the midpoint, this role pays about $200k versus about $188k for comparable roles.

Based on 240 similar postings.

Employer

About Goldman Sachs

Goldman Sachs is a leading global investment banking, securities, and investment management firm providing financial services to corporations, financial institutions, governments, and individuals.

Goldman Sachs currently has 134 open roles on FindRole.

Listed pay typically runs $137,000–$250,000 across 55 roles with salary data.

Most-posted roles

View all roles at Goldman Sachs

At a glance

TL;DR · Site Reliability Engineer, Global Banking & Markets, Vice President

Site Reliability Engineer, Global Banking & Markets, Vice President joins the Site Reliability Engineering team to ensure the availability, resilience, and performance of core business services for a global 24/7 trading operation. You will manage SLIs, SLOs, and error budgets while designing high-availability, multi-region, cloud-native services with integrated observability. The role involves leading incident response for latency-sensitive trade lifecycle systems, automating operational toil, and utilizing an AI-centric toolchain to accelerate root-cause analysis and production-quality automation. Key technical requirements include proficiency in Java 17+, Kubernetes, Docker, Terraform, and infrastructure-as-code. You will work with Apache Kafka, Spring Boot, gRPC, and Prometheus within a cloud environment involving GCP or AWS. The role addresses the critical challenge of maintaining high-throughput trade lifecycle management while navigating complex risk requirements and ensuring robust performance across front, middle, and back office functions.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Design and operate high-availability, multi-region, cloud-native services with integrated security and observability.
  • Establish and manage SLIs, SLOs, and error budgets to ensure the reliability of critical trading services.
  • Lead incident response for latency-sensitive systems by diagnosing issues and coordinating cross-functional teams.
  • Automate repetitive operational tasks and identify systemic risks before they impact production environments.
  • Orchestrate AI agents to accelerate root-cause analysis, code comprehension, and the creation of production-quality automation.
  • Develop event-driven architectures and optimized data paths for high-throughput trade lifecycle management.
  • Conduct risk-based activities including capacity planning, chaos engineering, and failover testing.
  • Translate post-incident findings into durable engineering improvements through blameless reviews.

What we're looking for

  • 8+ years of professional software or reliability engineering experience.
  • Proficiency in at least one major language, such as Java 17 (preferred).
  • Demonstrated risk acumen to identify and mitigate risks in a regulated financial environment.
  • Proven experience managing high-availability production environments including SLIs, SLOs, error budgets, and incident command.
  • Strong knowledge of cloud infrastructure (GCP, AWS), container orchestration (Kubernetes, Docker), and infrastructure-as-code.
  • Working knowledge of AI models and AI-assisted engineering tools to automate operations and manage codebases.
  • Experience building event-driven distributed systems using messaging platforms like Apache Kafka.
  • Strong skills in SDLC automation, observability discipline, and stakeholder coordination across technical and non-technical audiences.

More like this

Similar roles

Site Reliability Engineer, Lead

Booz Allen Hamilton

Chantilly, VA 50 days ago $99,000$225,000
Prometheus Grafana ELK Stack Linux AWS Python Terraform Terragrunt Kubernetes OpenTelemetry AWS CloudWatch AWS EKS Rancher Jenkins Git Docker Nessus JIRA Confluence SRE
8+ yrs exp