Site Reliability Engineer, Kafka

Apple Inc

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$142,300–$263,300 / yr
Posted
48 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $181k
This role $203k
$128k most similar roles pay here $278k

This role pays more than 67% of similar roles. Most pay $147,075–$214,000 — the shaded band above. At the midpoint, this role pays about $203k versus about $181k for comparable roles.

Based on 238 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Site Reliability Engineer, Kafka

Site Reliability Engineer - Kafka joins the Service Engineering Data Streaming SRE team to build and manage large-scale, fault-tolerant distributed systems. You will develop tools and automation for managing infrastructure, specifically focusing on next-generation Kafka platform services that provide low-latency data delivery. Daily responsibilities include contributing to Kafka deployment infrastructure, maintenance automation, control plane enhancements, monitoring, alerting, and performance engineering through profile-guided optimization. The role involves managing service lifecycles across bare metal, virtualized EC2 instances, and Kubernetes platforms while collaborating cross-functionally to define metrics and identify optimizations. Required technical skills include proficiency in Java, Go, or Python, along with experience in AWS, GCP, and Terraform for infrastructure as code. You will solve complex problems related to distributed systems, database storage engines, and multi-datacenter networking topologies to ensure reliable, scalable data infrastructure.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Develop and maintain tools and automation for managing large-scale distributed systems.
  • Build and optimize next-generation Kafka infrastructure and platform services.
  • Perform deep performance engineering, including design concepts and profile-guided optimizations.
  • Manage service lifecycles across bare metal, virtualized (EC2), and Kubernetes platforms.
  • Create monitoring, alerting tools, dashboards, and incident management runbooks.
  • Troubleshoot and perform deep dive analysis on distributed systems and database storage engines.
  • Design and manage multi-datacenter systems with a focus on high availability and scalability.
  • Develop infrastructure as code (IaC) using tools like Terraform for cloud and datacenter environments.

What we're looking for

  • 5 or more years of experience supporting internet-facing production services and distributed systems via deployments, On Call, and Incident Management.
  • 5 or more years of experience running large scale infrastructure with a heavy reliance on automation tooling.
  • 5 or more years of experience troubleshooting and performing deep dive analysis.
  • Proficiency in one or more programming languages including Java, Go (golang), or Python.
  • Operational experience managing services at scale on Kubernetes platforms.
  • Experience deploying and running services on Datacenter and Cloud architectures, including networking topologies and multi-datacenter systems.
  • Expertise developing and troubleshooting distributed systems and database storage engines.
  • Experience with AWS, GCP, and Infrastructure as Code (IaC) tools such as Terraform.

More like this

Similar roles

Site Reliability Engineer

Morgan Stanley

Alpharetta, GA 22 days ago
Python Shell Scripting Perl Ruby Java C# AWS Azure Jenkins Splunk DB2 Oracle Sybase Autosys Linux Unix Windows Agile Scrum Web Services MQ
5+ yrs exp

Site Reliability Engineer

Berkeley Research Group

Remote 77 days ago $130,000$160,000
Azure Kubernetes CI/CD GitHub Actions GitLab CI Golang Ruby Python AWS GCP Infrastructure as Code Datadog OpsGenie PagerDuty SRE Incident Management
5+ yrs exp Remote

Site Reliability Engineer

Balyasny Asset Management

Warsaw, Poland 78 days ago
Prometheus Grafana Loki Tempo OTEL Kubernetes Docker AWS Python Bash Go CI/CD DevOps SRE Agile
5+ yrs exp

Site Reliability Engineer

Booz Allen Hamilton

McLean, VA 9 days ago $86,800$198,000
AWS Kubernetes Terraform Ansible CI/CD Python Bash PowerShell Docker Prometheus Grafana Loki Elasticsearch Kibana GitLab GitHub CloudFormation OpenShift Jenkins REST JSON YAML XML Agile
6+ yrs exp

Site Reliability Engineer

Corepay

Brentwood, TN 2 days ago $157,000$190,000
AWS Azure Generative AI Git Agile Scrum Cloud Architecture Quality Engineering Source Control
6+ yrs exp