Staff DevOps Engineer

Anduril Industries

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Costa Mesa, CA
Salary
$191,000–$253,000 / yr
Posted
29 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $171k
This role $222k
$112k most similar roles pay here $268k

This role pays more than 85% of similar roles. Most pay $126,800–$214,500 — the shaded band above. At the midpoint, this role pays about $222k versus about $171k for comparable roles.

Based on 240 similar postings.

Employer

About Anduril Industries

Anduril Industries is a defense technology company that builds advanced hardware and software systems for national security, including autonomous drones, surveillance systems, and the Lattice AI command platform.

Anduril Industries currently has 1697 open roles on FindRole.

Listed pay typically runs $146,000–$194,000 across 1504 roles with salary data.

Most-posted roles

View all roles at Anduril Industries

At a glance

TL;DR · Staff DevOps Engineer

Staff DevOps Engineer The Staff DevOps Engineer joins the CorpTech Platform team to establish the reliability architecture for systems supporting business and manufacturing operations. This role focuses on designing and operating observability infrastructure, including metrics, distributed tracing, and structured logging, while managing deployment systems such as canary analysis and automated rollbacks. The engineer will define SLO frameworks, identify systemic risks, and develop reliability patterns specifically for AI-enabled systems to manage non-deterministic behavior. Key responsibilities include leading incident response for multi-system failures and setting production-readiness standards across the organization. Required technical skills include proficiency in Go, Python, or Rust, along with deep expertise in Kubernetes, cloud platforms like AWS, GCP, or Azure, and networking. The role addresses the challenge of scaling reliability as a core platform property rather than an individual effort within complex corporate infrastructure.

What does a DevOps Engineer earn?

Median $126800 from 89 postings across 26 companies.

See salary data

What you'll do

  • Design and manage the reliability architecture for production environments including observability, deployment systems, and incident management.
  • Build and operate an enterprise-scale observability platform featuring metrics, distributed tracing, and structured logging.
  • Develop release-safety mechanisms such as canary analysis, automated rollbacks, and progressive rollout systems.
  • Establish and govern SLO frameworks to make reliability measurable for engineering teams and leadership.
  • Identify systemic risks and implement automation to eliminate entire classes of failures rather than patching symptoms.
  • Define production-readiness standards and review processes to embed reliability into the software development lifecycle.
  • Lead incident response for complex multi-system failures and drive durable, systemic improvements from post-incident reviews.
  • Develop reliability patterns specifically for AI-enabled systems, including monitoring for model behavior drift and degradation.

What we're looking for

  • 10+ years of experience in site reliability engineering, production engineering, infrastructure engineering, or a related discipline at an architecture or platform-wide scope.
  • Demonstrated experience designing and owning reliability infrastructure such as observability platforms, deployment systems, or incident management tooling at scale.
  • Deep technical fluency in distributed systems, Kubernetes, cloud platforms (AWS, GCP, or Azure), networking, and storage.
  • Proficiency in systems programming languages including Go, Python, Rust, or equivalent for building production infrastructure and automation.
  • Demonstrated experience defining SRE standards and influencing adoption across engineering teams without formal authority.
  • Track record of leading incident response for complex failures and converting findings into systemic infrastructure improvements.
  • Degree in Computer Science, Information Systems, Engineering, or a related technical field, or equivalent practical experience.
  • U.S. Person status is required to access export controlled data.

More like this

Similar roles

Staff Site Reliability Engineer

Anduril Industries

Costa Mesa, CA 31 days ago $191,000$253,000
SRE Kubernetes AWS GCP Azure Go Python Rust Prometheus Grafana OpenTelemetry Datadog Distributed Systems Canary Analysis Feature Flagging Capacity Planning Cost Optimization
10+ yrs exp

Associate Staff DevOps Engineer

Abbott

Sunnyvale, CA +2 86 days ago $100,000$200,000
Azure Kubernetes Docker Terraform Bicep ARM Templates CI/CD GitHub Actions Jenkins Azure DevOps Python Bash PowerShell PostgreSQL Redis Firebase Prometheus Grafana DevSecOps Infrastructure as Code
10+ yrs exp

Associate Staff DevOps Engineer

Abbott

Sunnyvale, CA +2 21 days ago $100,000$200,000
Azure Kubernetes Docker Terraform Bicep ARM Templates CI/CD GitHub Actions Jenkins Azure DevOps Python Bash PowerShell PostgreSQL Redis Firebase Prometheus Grafana DevSecOps Infrastructure as Code
10+ yrs exp

DevOps Engineer

Booz Allen Hamilton

Dayton, OH 58 days ago $77,500$176,000
Kubernetes Docker AWS Azure Terraform Python Linux Shell Script CI/CD DevSecOps Containerization Jenkins Cloud-Native Monitoring Automation
5+ yrs exp

DevOps Engineer

Booz Allen Hamilton

Lexington, MA +3 17 days ago $77,600$176,000
CI/CD DevSecOps Terraform Python Shell Scripting PowerShell Ansible Kubernetes Docker GitLab CI/CD AWS Azure GCP OCI Agile Monitoring Regression Testing

DevOps Engineer

Booz Allen Hamilton

McLean, VA 60 days ago $77,600$176,000
AWS Kubernetes CI/CD GitOps Terraform Helm Argo CD Flux CD CloudFormation GitHub Actions Nexus Artifactory Istio Prometheus Grafana SQL Server S3 SQS SNS DynamoDB Keycloak