Senior High Performance Computing (HPC) Systems Engineer

SpaceX

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Starbase, TX
Posted
64 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $189k
$141k most similar roles pay here $234k

This listing doesn't post a salary. Most similar roles pay $153,100–$224,575.

Based on 240 similar postings.

Employer

About SpaceX

SpaceX designs, manufactures, and launches advanced rockets and spacecraft with the mission of enabling humans to become a multi-planetary species. It operates the Falcon 9, Falcon Heavy, and Starship launch vehicles, as well as the Starlink satellite internet constellation.

SpaceX currently has 680 open roles on FindRole.

Listed pay typically runs $130,000–$165,000 across 442 roles with salary data.

Most-posted roles

View all roles at SpaceX

At a glance

TL;DR · Senior High Performance Computing (HPC) Systems Engineer

As a Sr. High Performance Computing (HPC) Systems Engineer, you will join the HPC team to support SpaceX personnel and proprietary systems within a fast-paced environment. Your daily responsibilities include administering and managing HPC clusters, storage systems, and high-speed networks while providing application support across various engineering disciplines. You will install and integrate Linux-based compute clusters and create technical documentation for non-technical audiences. The role requires expertise in client and server hardware, enterprise networking, virtualization, security technologies, and Kubernetes. Preferred skills include experience with Bash or Python scripting, cluster resource managers like Slurm, PBS, or LSF; monitoring tools such as Prometheus, Grafana, and Nagios; and container engines like Docker, Podman, or Singularity. You will also work with GPU usage in compute clusters, Cuda, and automated configuration management software like Puppet or Ansible.

What you'll do

  • Administer and manage HPC clusters, storage systems, and high-speed networks.
  • Provide technical application support to employees across various engineering disciplines.
  • Install and integrate Linux-based compute clusters into the infrastructure.
  • Write instructional documentation and communicate complex technical ideas to non-technical users.
  • Automate recurring tasks and solve problems using scripting languages like Bash or Python.
  • Manage cluster resource managers such as Slurm, PBS, or LSF.
  • Deploy and maintain automated configuration management software like Puppet or Ansible.
  • Monitor and alert system performance using tools like Prometheus, Grafana, or Nagios.

What we're looking for

  • Bachelor's degree in computer science, engineering, math, or a scientific discipline and 5+ years of systems engineering experience.
  • 7+ years of professional experience building software in lieu of a degree.
  • 5+ years of hands-on experience with client/server hardware, management tools, enterprise networking, virtualization, and security technologies.
  • Experience with Kubernetes.
  • 5+ years of professional experience building, deploying, and troubleshooting Linux systems.
  • Experience with scripting languages like Bash or Python to automate tasks.
  • Familiarity with cluster resource managers (Slurm, PBS, LSF) and container engines (Docker, Podman, Singularity).
  • Eligibility for access to classified material up to TS/SCI with Polygraph and meeting ITAR requirements.

More like this

Similar roles

Site Reliability Engineer, HPC & Automation

SpaceX

Redmond, WA 72 days ago $125,000$150,000
HPC Python Bash Linux Docker Kubernetes Terraform Ansible Puppet Prometheus Grafana CI/CD MySQL PostgreSQL SQLite TCP/IP Slurm LSF REST API NFS Cadence Synopsys
2+ yrs exp

Site Reliability Engineer

SpaceX

Hawthorne, CA 59 days ago $125,000$150,000
High Performance Computing Linux Windows Server Infiniband Python Bash Puppet Ansible Kubernetes Docker Networking Virtualization CFD FEA ANSYS StarCCM+
1+ yrs exp

System Software Engineer, HPC Performance

Nvidia

Champaign, IL +2 17 days ago $152,000$241,500
C C++ Python HPC Machine Learning Deep Learning Artificial Intelligence x86 ARM Linux Windows macOS Profiling Tools Cloud Computing
5+ yrs exp

HPC Systems Engineer, Modeling & Simulation

Anduril Industries

Costa Mesa, CA 92 days ago $132,000$198,000
HPC Linux Unix Python Bash Slurm MPI OpenMP CMake NFS NAS TCP/IP ParaView VisIt Cubit CTH ALE3D Sierra Data Workflows
5+ yrs exp

AI Systems Engineer, HPC

Amd

San Jose, CA 24 days ago $173,600$260,400
GPU HPC Kubernetes Python SLURM RoCEv2 KVM Ubuntu Shell Ansible Saltstack Terraform Prometheus Grafana Distributed ML LLMs 400G Networking