Senior Site Reliability Engineer
Okta Inc
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 54% of similar roles. Most pay $149,580–$214,625 — the shaded band above. At the midpoint, this role pays about $192k versus about $182k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 896 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.
Most-posted roles
At a glance
As a Senior Site Reliability Engineer, AIOPs, you will join the team building an AI Data Center AIOps platform that transforms high-volume telemetry into reliable insights and automation for GPU fleets. You will be responsible for the platform's uptime, performance, data integrity, and safe change management rather than the compute cluster itself. Your daily work involves managing SLOs/SLIs, incident response, and postmortems for telemetry ingestion, processing, storage, and APIs. You will manage Kubernetes deployments using Helm and Terraform to ensure scalable environments while automating recurring checks and building comprehensive runbooks. The role requires expertise in Python, Bash, CI/CD pipelines, and infrastructure-as-code. You will also navigate complex distributed systems involving Kafka, Pulsar, Flink, Spark, ClickHouse, Elastic, and TSDBs to solve challenges related to backpressure, hotspots, and failure domains within the observability domain.
What does a Site Reliability Engineer earn in California?
Median $214000 from 54 postings across 15 companies.
Skills
What you'll do
What we're looking for
More like this
Okta Inc
The Federal Reserve
Autodesk
Autodesk
Salesforce
MongoDB