Compute Site Reliability Engineering Manager

Apple Inc

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$225,600–$338,400 / yr
Posted
119 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $186k
This role $282k
$127k most similar roles pay here $361k

This role pays more than 96% of similar roles. Most pay $155,437–$216,062 — the shaded band above. At the midpoint, this role pays about $282k versus about $186k for comparable roles.

Based on 238 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Compute Site Reliability Engineering Manager

The ASE Compute - Site Reliability Engineering (SRE) Manager leads a team responsible for the reliability, availability, and performance of mission-critical cloud platform services. This hands-on leadership role involves mentoring engineers, driving automation to reduce operational toil, and managing on-call rotations while establishing SRE practices like SLOs, error budgets, and capacity planning. The manager will partner with software and architecture teams to influence system design for scalability and operability. The role focuses on the private cloud infrastructure that powers services through bare-metal Kubernetes clusters and virtualized environments. Key technical requirements include experience with multi-tenant Kubernetes environments, configuration management tools like Puppet or Ansible, and troubleshooting across network, OS, and container runtimes. Preferred skills include familiarity with Java, Go, Python, Prometheus, Thanos, Splunk, and managing infrastructure as an internal managed service with defined SLAs.

What you'll do

  • Lead and mentor a team of Site Reliability Engineers managing large-scale Kubernetes and compute infrastructure.
  • Own the reliability, availability, and performance of mission-critical cloud platform services.
  • Manage incident response, post-incident reviews, and systemic improvements to reduce operational toil.
  • Influence system design for scalability and operability by partnering with software and architecture teams.
  • Establish and refine SRE practices including SLOs, error budgets, capacity planning, and change management.
  • Drive automation to eliminate manual processes through custom tooling and self-service capabilities.
  • Manage on-call rotations to ensure sustainable and well-supported operational coverage.

What we're looking for

  • 5+ years of engineering management experience leading infrastructure or SRE teams.
  • Deep experience operating large-scale, multi-tenant Kubernetes environments in production.
  • Strong systems background troubleshooting across network, OS, container runtime, and application layers.
  • Experience with configuration management at scale using Puppet, Ansible, or equivalent tools.
  • Track record of building high-performing teams through coaching and clear expectations.
  • Demonstrated ability to drive cross-functional initiatives to completion.
  • Strong written and verbal communication skills.
  • CNCF Certified Kubernetes Administrator (CKA) or equivalent hands-on certification preferred.

More like this

Similar roles

Senior Site Reliability Engineer

Apple Inc

Cupertino, CA 76 days ago $150,400$277,600
Python Go Java Shell Scripting Kubernetes CI/CD Distributed Systems Observability SLOs SLIs Capacity Planning Cloud Infrastructure Automation SRE
5+ yrs exp

Senior Site Reliability Engineer

Apple Inc

San Francisco, CA 148 days ago $184,700$277,600
SRE Site Reliability Engineering Python Go Java Kubernetes Linux micro-services Distributed Systems Automation Networking Security Encryption Monitoring Alerting Capacity Planning Disaster Recovery
5+ yrs exp

Senior Site Reliability Engineer

Apple Inc

Cupertino, CA 158 days ago $184,700$277,600
SRE Python Go Java Kubernetes Linux Distributed Systems micro-services Automation Networking Security Encryption Monitoring Alerting Capacity Planning Disaster Recovery
5+ yrs exp

Site Reliability Engineer, AI Platform & Cloud

Morgan Stanley

Alpharetta, GA 143 days ago
Kubernetes AWS Azure Python Go Java Docker Terraform Helm CloudFormation Ansible Prometheus Grafana ELK Datadog Kafka Spark Flink SQL Redis Snowflake REST Infrastructure-as-Code SRE ML Ops GenAI
5+ yrs exp

Senior SRE Software Engineer

Apple Inc

San Francisco, CA 38 days ago $184,700$324,800
Kubernetes Go SRE DevOps Prometheus Thanos Splunk Puppet Ansible AWS GCP Azure Bare-metal Container Runtime
8+ yrs exp

SRE Software Engineer

Apple Inc

Austin, TX 49 days ago
Kubernetes Python Go Linux Docker CI/CD Git Ansible Puppet AWS GCP Azure Prometheus Thanos Splunk TCP/IP DNS DHCP Site Reliability Engineering