Service Reliability Engineer, G&A Solutions Engineering

Apple Inc

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Austin, TX
Posted
11 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $179k
$133k $229k
below market most similar roles pay here above market

This listing doesn't post a salary. Most similar roles pay $142,500–$216,250.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 3552 open roles on FindRole.

Listed pay typically runs $166,600–$277,600 across 2742 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Service Reliability Engineer, G&A Solutions Engineering

The Service Reliability Engineer, G&A Solutions Engineering joins the General and Administrative Solutions Engineering team to maintain the health, stability, and efficiency of global, mission-critical production systems. This role involves proactively monitoring service performance, identifying bottlenecks, and leading incident response efforts including root cause analysis. You will develop automation strategies to streamline operational tasks, define service level indicators, and collaborate with engineers, data engineers, and network specialists to ensure new services are designed for operational perfection. Required skills include proficiency in Python, Java, or Go, along with Bash or PowerShell scripting. You will utilize cloud platforms like AWS, Azure, or GCP, and cloud-native technologies such as Kubernetes and Docker. Experience with Prometheus, Grafana, Splunk, Datadog, and configuration management tools like Ansible is also required.

What you'll do

  • Monitor service performance to identify bottlenecks and implement solutions for improved efficiency and resilience.
  • Lead incident response efforts and conduct thorough root cause analysis to resolve production issues.
  • Develop and implement automation strategies to streamline operational tasks and reduce manual intervention.
  • Apply SRE principles to maintain highly reliable and scalable service infrastructure.
  • Partner with development teams to ensure new services include best practices for monitoring and scalability.
  • Create and maintain comprehensive documentation, including run-books and service level objectives.
  • Participate in on-call rotations to provide 24/7 support for critical services.
  • Define and supervise key service level indicators to measure and improve overall service reliability.

What we're looking for

  • 4+ years of experience in Site Reliability Engineering, production support, or a related role supporting large-scale enterprise services.
  • Bachelor's degree in Computer Science or work-related equivalent experience.
  • Strong proficiency in at least one programming language such as Python, Java, or Go.
  • Proficiency in scripting languages such as Bash or PowerShell.
  • Experience with cloud platforms like AWS, Azure, or GCP and cloud-native technologies like Kubernetes and Docker.
  • Hands-on experience with monitoring and alerting tools such as Prometheus, Grafana, Splunk, or Datadog.
  • Familiarity with CI/CD pipelines, DevOps practices, and configuration management tools like Ansible, Chef, or Puppet (preferred).
  • Experience with database technologies, ITIL frameworks, Linux/Unix administration, and vibe coding (preferred).

More like this

Similar roles

Senior Site Reliability Engineer, Customer Systems

Apple Inc

Austin, TX 144 days ago
Kubernetes Helm Python Ansible Shell Scripting CI/CD Splunk Grafana Prometheus Alertmanager Cassandra MongoDB Couchbase AWS S3 ArgoCD GitOps Java DNS TCP HTTP/HTTPS Infrastructure as Code
5+ yrs exp

Senior Site Reliability Engineer, Customer Systems

Apple Inc

Austin, TX 23 days ago
Kubernetes Helm Python Ansible Shell Scripting CI/CD Splunk Grafana Prometheus Alertmanager Cassandra MongoDB Couchbase AWS S3 ArgoCD GitOps Java DNS TCP HTTP/HTTPS Infrastructure as Code
5+ yrs exp

Site Reliability Engineer, Customer Systems

Apple Inc

Sunnyvale, CA 144 days ago $150,400–$225,300
Kubernetes Helm Python Ansible Shell Scripting CI/CD Splunk Grafana Prometheus Alertmanager ArgoCD GitOps Java DNS TCP HTTP/HTTPS Infrastructure as Code GenAI

Site Reliability Engineer, Customer Systems

Apple Inc

Sunnyvale, CA 127 days ago $150,400–$225,300
Kubernetes Helm Python Ansible Shell Scripting CI/CD Splunk Grafana Prometheus Alertmanager ArgoCD GitOps Java DNS TCP HTTP/HTTPS Infrastructure as Code GenAI

Senior Site Reliability Engineer, Infrastructure Services

Apple Inc

San Francisco, CA 16 days ago $184,700–$324,800
Python Go Java Bash CI/CD Jenkins GitHub Actions GitLab Prometheus Grafana Splunk Datadog Docker Kubernetes AWS GCP Oracle MongoDB Kafka RabbitMQ Linux TLS/SSL DNS OAuth SAML SSO
3+ yrs exp