Lead Site Reliability Engineer, Vice President

Morgan Stanley

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
New York, NY
Salary
$150,000–$190,000 / yr
Posted
44 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $185k
This role $170k
$134k most similar roles pay here $239k

This role pays less than 67% of similar roles. Most pay $155,437–$214,000 — the shaded band above. At the midpoint, this role pays about $170k versus about $185k for comparable roles.

Based on 238 similar postings.

Employer

About Morgan Stanley

Morgan Stanley is a global financial services firm providing investment banking, securities, wealth management, and investment management services to corporations, governments, institutions, and individuals. Industry: Investment Banking & Financial Services

Morgan Stanley currently has 38 open roles on FindRole.

Listed pay typically runs $150,000–$210,000 across 31 roles with salary data.

Most-posted roles

View all roles at Morgan Stanley

At a glance

TL;DR · Lead Site Reliability Engineer, Vice President

Lead Site Reliability Engineer, Vice President joins the WM Product Technology team to ensure high availability and reliability for production systems. This role focuses on minimizing outages by automating deployments, managing the full service lifecycle from design through capacity planning, and scaling systems via automation. The engineer will monitor system health, perform root cause analysis for incidents, troubleshoot infrastructure issues, and develop tools to streamline production management. Key technical requirements include proficiency in Python, Perl, Shell, Ruby, Java, or C#, along with experience in DB2, Sybase, or Oracle databases. Candidates must be skilled in Autosys, Jenkins, Splunk, and Sockeye while possessing expertise in Linux/Unix environments, cloud-based deployments in Azure and AWS, and containerized environments. The role addresses the critical need for stable production systems within a complex financial services technology infrastructure.

What you'll do

  • Monitor and measure application availability, latency, and system health while managing costs and reducing operational toil.
  • Automate deployments and routine operational tasks using scripting languages like Python, Perl, or Shell.
  • Troubleshoot infrastructure issues, review log files, and maintain a comprehensive knowledge base of resolutions.
  • Collaborate with development teams to build tools and utilities for production management and system stability.
  • Perform root cause analysis for outages and manage production requests in high-pressure environments.
  • Test and tune network, hardware, and software configurations to maximize overall performance.
  • Manage both production and non-production environments while participating in a 24/7 on-call rotation.
  • Identify system trends and risks to lead initiatives that improve reliability and scalability.

What we're looking for

  • Minimum of 10 years of industry experience, preferably within the financial IT community.
  • Minimum of 10 years of hands-on experience in designing, developing, and implementing technical solutions or deep technical support.
  • Proficiency in scripting languages including Python, Perl, Shell, Ruby, Java, or C#.
  • Strong database skills with DB2, Sybase, or Oracle.
  • Experience with batch scheduling software such as Autosys.
  • Expertise in CI/CD, cloud-driven development (Azure and AWS), and containerized environments.
  • Proficiency with monitoring tools like Splunk, IP Soft, and Sockeye.
  • Experience with Agile methodologies and automating deployments using Jenkins, Train, or Windeploy.

More like this

Similar roles

Lead Site Reliability Engineer

Morgan Stanley

Alpharetta, GA 17 days ago $125,000$175,000
SRE Python Shell Scripting Perl AWS GCP Azure DB2 Oracle Sybase Autosys Prometheus Grafana Splunk Kibana Chaos Engineering Adobe Experience Cloud
10+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 45 days ago
Site Reliability Engineering Python Java Spring Boot .Net AI CI/CD Container Orchestration AWS Observability Monitoring Telemetry Networking System Architecture SDLC
5+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Jersey City, NJ 43 days ago
Site Reliability Engineering Python Java Spring Boot .Net Kubernetes AWS Google Cloud CI/CD Observability Monitoring Telemetry Networking Infrastructure Optimization FinOps Disaster Recovery Capacity Management
5+ yrs exp

Site Reliability Engineer

Morgan Stanley

Alpharetta, GA 22 days ago
Python Shell Scripting Perl Ruby Java C# AWS Azure Jenkins Splunk DB2 Oracle Sybase Autosys Linux Unix Windows Agile Scrum Web Services MQ
5+ yrs exp

Lead Principal Site Reliability Engineer

Oracle

Nashville, TN 56 days ago $96,300$264,100
Site Reliability Engineering Kubernetes Docker Terraform Ansible Chef Puppet Python Go Java JavaScript Bash Oracle Cloud Infrastructure Microsoft Azure Google Cloud Platform infrastructure-as-code Chaos Engineering
6+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Palo Alto, CA 59 days ago
Site Reliability Engineering Java Go Python Terraform Kubernetes Docker CI/CD GitOps Grafana Prometheus Dynatrace Datadog Splunk Kafka RabbitMQ SQS Neo4j Pinecone Weaviate Chroma LangChain LangGraph AutoGen CrewAI GitHub Copilot Fluentd Logstash Vector RESTful APIs RAG TensorFlow PyTorch scikit-learn Hadoop Spark Flink MongoDB Cassandra DynamoDB InfluxDB TimescaleDB AWS Azure GCP
5+ yrs exp