Site Reliability Engineer, Apple Data Platform

Apple Inc

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Austin, TX
Posted
43 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $202k
$139k most similar roles pay here $263k

This listing doesn't post a salary. Most similar roles pay $172,000–$231,150.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Site Reliability Engineer, Apple Data Platform

Site Reliability Engineer, Apple Data Platform joins the Data Platform Site Reliability Engineering team to manage infrastructure and applications on bare-metal and cloud computing platforms. This role involves delivering data processing, governance, and storage for global products while ensuring high availability and performance across geographically distributed data centers. The engineer will collaborate with partner teams to implement consistent incident management processes, automate deployments, and establish user journey based SLOs derived from observability metrics. Key technical requirements include proficiency in Go, Python, or Java, along with experience in Kubernetes, Linux operating systems, containers, and standard networking protocols. Candidates should possess expertise in managing distributed systems, capacity planning, and disaster recovery. Preferred skills include experience with Flink, Hive, Hadoop/HDFS, Trino, Druid, and cloud providers such as AWS, GCP, and Ali Cloud to manage large-scale data analytics.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Manage infrastructure and applications on bare-metal and cloud computing platforms for global data processing.
  • Maintain high availability, performance, and scalability for distributed systems across geographically dispersed data centers.
  • Develop and release code in Go, Python, or Java to manage configuration and software delivery.
  • Implement automated deployment processes and eliminate manual toil through automation.
  • Monitor and improve system performance using observability metrics and user journey-based SLOs.
  • Execute incident management processes and perform troubleshooting for large-scale production applications.
  • Optimize and troubleshoot open-source data analytics and governance technologies like Flink, Hive, or Trino.
  • Perform capacity planning and disaster recovery testing for distributed systems on Kubernetes and public clouds.

What we're looking for

  • BS/MS in Computer Science or equivalent degree.
  • 5+ years of software development or production operations experience in a large-scale environment.
  • Proficiency in authoring and releasing code in Go, Python, or Java using configuration management and delivery platforms.
  • Experience operating production applications at scale including performance testing, HA, disaster recovery, and capacity planning.
  • Experience managing distributed systems on internal and public cloud infrastructure, specifically Kubernetes.
  • Understanding of Linux OS, containers, virtualization, and standard networking protocols.
  • Proficiency in troubleshooting and tuning open source data analytics or governance technologies like Flink, Hive, Hadoop/HDFS, Trino, or Druid.
  • Proficiency in managing applications and infrastructure on AWS, GCP, and Ali Cloud.

More like this

Similar roles

Senior Site Reliability Engineer

Apple Inc

San Francisco, CA 148 days ago $184,700$277,600
SRE Site Reliability Engineering Python Go Java Kubernetes Linux micro-services Distributed Systems Automation Networking Security Encryption Monitoring Alerting Capacity Planning Disaster Recovery
5+ yrs exp