Site Reliability Engineer, Data Center Infrastructure

SpaceX

Confirmed live today High trust

Quick summary

Work type
On-site
Location
Bastrop, TXHawthorne, CARedmond, WACape Canaveral, FLStarbase, TX
Posted
3 days ago
Freshness
Confirmed live today

Market check

Salary context

How this pay compares to similar roles

Similar $182k
$135k $237k
below market most similar roles pay here above market

This listing doesn't post a salary. Most similar roles pay $145,000–$219,237.

Based on 240 similar postings.

Employer

About SpaceX

SpaceX designs, manufactures, and launches advanced rockets and spacecraft with the mission of enabling humans to become a multi-planetary species. It operates the Falcon 9, Falcon Heavy, and Starship launch vehicles, as well as the Starlink satellite internet constellation.

SpaceX currently has 796 open roles on FindRole.

Listed pay typically runs $130,000–$165,000 across 555 roles with salary data.

Most-posted roles

View all roles at SpaceX

At a glance

TL;DR · Site Reliability Engineer, Data Center Infrastructure

The Site Reliability Engineer, Data Center Infrastructure joins the application software team to manage the compute, storage, and networking infrastructure supporting manufacturing systems for Starship, Starlink, Starshield, and Terafab. The role involves deploying, upgrading, and scaling production systems while ensuring factory uptime and throughput. Key responsibilities include managing infrastructure as code, utilizing observability for platform health, performing capacity planning, and conducting proactive maintenance to reduce toil. The candidate must possess software engineering fundamentals, Linux operating system experience, and at least one year of software development experience. Preferred technical skills include Terraform, Ansible, Puppet, Docker, Kubernetes, vSphere, QEMU, KVM, Postgres, and Clickhouse. This position solves critical reliability and scalability problems for mission-critical manufacturing environments, requiring the ability to translate high-level requirements into stable, maintainable implementations.

What does a Site Reliability Engineer earn in California?

Median $214000 from 72 postings across 20 companies.

See salary data

What you'll do

  • Deploy, upgrade, operate, and scale compute, storage, and networking for manufacturing systems.
  • Manage infrastructure as code and use observability tools to monitor platform health.
  • Design systems for reliability and scale while identifying and removing performance bottlenecks.
  • Perform proactive maintenance including capacity planning and lifecycle management to reduce toil.
  • Improve the full infrastructure lifecycle from initial design through deployment and continuous refinement.
  • Practice sustainable incident response and conduct blameless postmortems.
  • Provide high-quality technical support to manufacturing and engineering users.
  • Participate in on-call rotations and travel to sites for deployments and incident response.

What we're looking for

  • Bachelor’s degree in computer science, information systems, or an engineering discipline; OR 3+ years of professional experience in SRE or DevOps in lieu of a degree.
  • 1+ years of software development experience.
  • Experience with Linux operating systems.
  • Experience with compute, storage, and/or networking infrastructure in production (preferred).
  • Experience with Infrastructure as Code (Terraform, Ansible, Puppet, or similar) (preferred).
  • Experience with containers and virtualization (Docker, Kubernetes, vSphere, QEMU, KVM, etc.) (preferred).
  • Experience with databases and data modeling (Postgres, Clickhouse, etc.) (preferred).
  • Must be a U.S. citizen, national, lawful permanent resident, refugee, asylee, or eligible for required Department of State authorizations.

More like this

Similar roles

Site Reliability Engineer, HPC & Automation

SpaceX

Redmond, WA 102 days ago $125,000–$150,000
High Performance Computing Python Bash Linux Docker Kubernetes Terraform Ansible Puppet CI/CD Jenkins Prometheus Grafana MySQL PostgreSQL SQLite TCP/IP REST API NFS Cadence Synopsys Ansys Keysight Siemens
2+ yrs exp

Site Reliability Engineer, Application Software

SpaceX

Hawthorne, CA 80 days ago $125,000–$160,000
Python Linux Docker Kubernetes Terraform Ansible MySQL ClickHouse JavaScript C# C++ Bazel Buck Make Infrastructure as Code Observability vSphere QEMU KVM