Service Reliability Engineer

Apple Inc

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$142,300–$263,300 / yr
Posted
49 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $180k
This role $203k
$128k most similar roles pay here $278k

This role pays more than 66% of similar roles. Most pay $145,000–$215,450 — the shaded band above. At the midpoint, this role pays about $203k versus about $180k for comparable roles.

Based on 238 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Service Reliability Engineer

The Service Reliability Engineer (SRE) joins the Apple Services Engineering team to ensure the performance, stability, and availability of multi-tiered systems supporting global products like Apple Music and the App Store. This role involves managing jobs and applications on bare-metal and cloud computing platforms while processing exabytes of data for large-scale analytics. Key responsibilities include configuring and tuning systems, building automation for self-healing capabilities, creating monitoring tools for low-latency applications, and troubleshooting complex network and performance issues across production and development environments. The candidate will work with Java-based applications, Spark, Flink, Hadoop, Kubernetes, and AWS infrastructure. This position addresses the technical challenge of maintaining high-performance data processing systems within a dynamic environment, requiring expertise in SRE principles to ensure reliable service delivery for diverse media content across various global platforms.

What you'll do

  • Support Java-based applications and Spark/Flink jobs across Baremetal, AWS, and Kubernetes platforms.
  • Configure and tune multi-tiered systems to optimize application performance, stability, and availability.
  • Assess application requirements to determine the appropriate services and topology for cloud and on-premise environments.
  • Build automation to create self-healing systems.
  • Develop tools to monitor high-performance systems and alert low-latency applications.
  • Troubleshoot complex issues involving core networking, system performance, and specific application errors.
  • Monitor production, staging, test, and development environments across a diverse range of applications.

What we're looking for

  • At least 5 years of experience in a Site Reliability Engineering (SRE) or DevOps role.
  • A BS degree in computer science with 5+ years of experience, an MS degree with 3+ years of experience, or equivalent.
  • At least 5 years of experience running services in a large-scale *nix environment.
  • Understanding of SRE principles and goals along with prior on-call experience.
  • Extensive experience managing applications on AWS and Kubernetes.
  • Deep understanding and experience in Hadoop, Spark, Flink, Kubernetes, or AWS.
  • Experience supporting Java applications and Big Data technologies.
  • Strong analytical problem solving, communication skills, and ability to work with geographically distributed teams.

More like this

Similar roles

SRE Software Engineer

Apple Inc

Austin, TX 49 days ago
Kubernetes Python Go Linux Docker CI/CD Git Ansible Puppet AWS GCP Azure Prometheus Thanos Splunk TCP/IP DNS DHCP Site Reliability Engineering

SRE Software Engineer

Apple Inc

Austin, TX 44 days ago
Kubernetes Python Go Linux Docker CI/CD Git Ansible Puppet AWS GCP Azure Prometheus Thanos Splunk TCP/IP DNS DHCP Site Reliability Engineering

SRE Software Engineer

Apple Inc

Austin, TX 13 days ago
Kubernetes Python Go Linux Docker CI/CD Git Ansible Puppet AWS GCP Azure Prometheus Thanos Splunk TCP/IP DNS DHCP Site Reliability Engineering

Site Reliability Engineer, AI Platform & Cloud

Morgan Stanley

Alpharetta, GA 143 days ago
Kubernetes AWS Azure Python Go Java Docker Terraform Helm CloudFormation Ansible Prometheus Grafana ELK Datadog Kafka Spark Flink SQL Redis Snowflake REST Infrastructure-as-Code SRE ML Ops GenAI
5+ yrs exp

Senior Site Reliability Engineer

CVS Health

Woonsocket, RI +1 14 days ago $92,700$203,940
SRE DevOps Kubernetes OpenShift Docker CI/CD Splunk Dynatrace Datadog Prometheus Grafana Python Java AWS Microsoft Azure Google Cloud Rancher GitHub BitBucket Jenkins Microservices web API’s Apigee
5+ yrs exp Hybrid

Senior Manager Site Reliability Engineer

The Walt Disney Company

Remote 16 days ago $175,000$215,000
SRE AWS GCP Azure Kubernetes Terraform Ansible Harness GitLab CloudFormation CI/CD Observability Automation Serverless DevOps
10+ yrs exp Remote