Principal Site Reliability Engineer
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 72% of similar roles. Most pay $159,875–$224,250 — the shaded band above. At the midpoint, this role pays about $219k versus about $192k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 896 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.
Most-posted roles
At a glance
Staff Site Reliability Engineer - AI Platform Runtime joins the SRE team to lead technical strategy and roadmaps for large-scale initiatives improving reliability, scalability, and developer productivity across enterprise systems. The role involves designing and building resilient distributed systems that power next-generation AI-driven products while architecting AI Agents and Skills to accelerate platform operations. Day-to-day responsibilities include driving automation and observability improvements using metrics and analytics, troubleshooting complex systems, and mentoring engineers to foster technical excellence. Candidates must possess expertise in Python, TypeScript, JavaScript, or Go, alongside experience with infrastructure-as-code tools like AWS CDK, CloudFormation, Terraform, or CrossPlane. The role requires deep knowledge of Kubernetes, networking, and public cloud services such as AWS, Azure, or GCP, while implementing OpenTelemetry for observability at scale to ensure high availability for AI infrastructure.
What does a Site Reliability Engineer earn in California?
Median $214000 from 54 postings across 15 companies.
Skills
What you'll do
What we're looking for
More like this
Nvidia
TransUnion
CME Group
MongoDB
Circle
Anduril Industries