Senior Engineering Manager, Site Reliability

Upstart

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
Canada
Salary
$195,300–$270,400 / yr
Posted
55 days ago
Freshness
Confirmed live yesterday
Closes
Dec 8, 2026

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $190k
This role $233k
$134k most similar roles pay here $285k

This role pays more than 87% of similar roles. Most pay $160,259–$219,125 — the shaded band above. At the midpoint, this role pays about $233k versus about $190k for comparable roles.

Based on 238 similar postings.

Employer

About Upstart

Upstart is an AI lending platform that partners with banks and credit unions to expand access to affordable credit using non-traditional variables.

Upstart currently has 73 open roles on FindRole.

Listed pay typically runs $166,900–$230,450 across 72 roles with salary data.

Most-posted roles

View all roles at Upstart

At a glance

TL;DR · Senior Engineering Manager, Site Reliability

Senior Engineering Manager, Site Reliability leads a team dedicated to ensuring the reliability, observability, and resilience of systems at scale. In this role, you will manage a team focused on incident management, operational readiness, and reliability engineering while translating high-level strategy into actionable roadmaps and measurable outcomes. You will build out automated safeguards, improve postmortem quality, and establish standards for service level objectives to reduce manual toil and detection gaps. The ideal candidate possesses deep technical expertise in distributed systems, cloud infrastructure, and production operations. Key technologies and tools mentioned include Kubernetes, AWS, and observability platforms such as Datadog, Grafana, Prometheus, or OpenTelemetry. You will solve complex problems regarding how the organization manages high-severity incidents and ensures that engineering teams can operate services safely while maintaining a robust, scalable infrastructure for the company's products.

What does a Engineering Manager earn in Remote?

Median $235250 from 43 postings across 20 companies.

See salary data

What you'll do

  • Manage and develop a team focused on incident management, observability, and reliability engineering.
  • Define the SRE function's charter, roadmap, and measurable outcomes to improve system reliability.
  • Lead high-severity incident response and evolve the program to improve detection and recovery.
  • Improve postmortem quality to ensure incident learnings result in durable engineering improvements.
  • Enhance production observability by improving the quality and trustworthiness of metrics, logs, and traces.
  • Establish scalable operational readiness standards for new services and major architectural changes.
  • Partner with product and infrastructure teams to embed reliability into standard engineering workflows.
  • Drive systemic improvements through automation, failure testing, and reduced operational toil.

What we're looking for

  • Must have at least 5 years of reliability engineering management experience.
  • Must have at least 7 years of experience in software, site reliability, infrastructure, or platform engineering.
  • Must have significant hands-on experience in Site Reliability Engineering or Production Engineering roles.
  • Must have direct experience managing an SRE or production engineering function including strategy and roadmap ownership.
  • Must possess strong technical depth in distributed systems, cloud infrastructure, observability, and production operations.
  • Must have experience leading high severity incident response and improving incident management practices at scale.
  • Must demonstrate the ability to hire, develop, and retain high performing engineers and leaders.
  • Must possess strong cross-functional leadership skills to translate complex data into clear decisions and alignment.

More like this

Similar roles

Senior Manager, Site Reliability Engineering

Intuit

Mountain View, CA 21 days ago $222,000$300,500
AWS Kubernetes Terraform CloudFormation SRE AIOps infrastructure-as-code Prometheus Grafana Datadog Splunk PagerDuty EC2 EKS ECS RDS DynamoDB CloudWatch Chaos Engineering Incident Management
8+ yrs exp

Senior Site Reliability Engineer

Salesforce

Remote (San Francisco, CA) 58 days ago $148,500$223,900
SRE Python Go Docker Kubernetes CI/CD Prometheus Grafana ELK Splunk Datadog Temporal Airflow Argo Workflows AWS GCP Linux Unix LLM Prompt Engineering
5+ yrs exp Remote

Senior Site Reliability Engineer

Autodesk

Remote (ID) +1 28 days ago $117,000$209,330
Site Reliability Engineering AWS Kubernetes Python Go Java Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace Bash PowerShell FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp Remote

Senior Software Engineer

Robinhood

New York, NY 94 days ago $196,000$230,000
Reliability Engineering Observability Prometheus Grafana OpenTelemetry Distributed Systems Incident Management Infrastructure Fault-tolerant Architecture Capacity Planning Failover Strategies MTTD MTTR Postmortems
5+ yrs exp Hybrid

Senior Manager, Site Reliability Engineering

Oracle

Nashville, TN 30 days ago $121,500$264,100
AIOps Automation Business Intelligence Data Center Operations Incident Management Network Operations Network Routing Structured Cabling Deployment Telecommunications Networking
3+ yrs exp