Senior Staff Site Reliability Operations

Nvidia

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$184,000–$264,500 / yr
Posted
4 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $190k
This role $224k
$137k most similar roles pay here $278k

This role pays more than 79% of similar roles. Most pay $160,259–$220,081 — the shaded band above. At the midpoint, this role pays about $224k versus about $190k for comparable roles.

Based on 238 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 929 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 915 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Staff Site Reliability Operations

Senior Staff Site Reliability Operations serves as a senior technical individual contributor and lead for site reliability and support. This role manages day-to-day operations including incident management, service quality, asset inventory, and vulnerability remediation while acting as the final escalation point for complex issues. The position involves leading site-level projects, mentoring engineers, and building automation to eliminate recurring problems. Key technical domains include Active Directory, hybrid Entra ID, Exchange hybrid mail flow, and various compute platforms including Windows, Linux, and macOS. You will manage infrastructure involving virtualization, storage, and datacenter hardware while utilizing tools like Intune, Autopilot, and ServiceNow. Required skills include scripting in PowerShell, Python, or Bash to drive automation. The role addresses critical infrastructure challenges across identity, messaging, and endpoint management within a complex enterprise environment to ensure high availability and security.

What you'll do

  • Serve as the final point of escalation for complex issues involving Active Directory, Exchange, database platforms, and compute infrastructure.
  • Manage day-to-day site operations including incident response, service level agreements (SLAs), and asset inventory management.
  • Drive root cause analysis to eliminate recurring technical problems rather than performing simple break-fix actions.
  • Lead the site support team by setting technical standards, mentoring engineers, and maintaining the knowledge base.
  • Manage endpoint compliance, vulnerability remediation, patch management, and security hardening in partnership with InfoSec.
  • Develop automation scripts in PowerShell, Python, or Bash to improve diagnostics, reporting, and system health checks.
  • Represent local site requirements and priorities in regional and global IT architecture and policy forums.
  • Communicate technical risks and incident updates to executive leadership and internal stakeholders.

What we're looking for

  • 12+ years of experience in enterprise support engineering, infrastructure, or end user services.
  • 5+ years of experience in a senior, lead, or escalation-tier role in a multi-site environment.
  • Deep hands-on expertise in Active Directory, hybrid Entra ID, Exchange hybrid, Windows/Linux servers, virtualization, and datacenter hardware.
  • Experience with enterprise endpoint management (Intune, Autopilot, MECM, Jamf), M365 ecosystem, and vulnerability remediation.
  • Proficiency in networking fundamentals including DNS, DHCP, VLAN, wireless, firewall policy, and switch-level troubleshooting.
  • Ability to script and automate using Python, PowerShell, or Bash for diagnostics and reporting.
  • Demonstrated technical leadership, executive-level communication skills, and experience with ServiceNow or similar ITSM tools.
  • Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent experience.

More like this

Similar roles

Senior Staff Site Reliability Operations Technical Lead

Nvidia

Durham, NC 4 days ago $184,000$264,500
Active Directory Entra ID Exchange Microsoft 365 Intune PowerShell Python Bash Linux Virtualization MECM SCCM Jamf Autopilot DNS DHCP VLAN ServiceNow ITSM Root Cause Analysis
10+ yrs exp

Senior Site Reliability Engineer

Oracle

Nashville, TN 21 days ago $81,100$187,000
OCI AWS Azure GCP Terraform Chef Ansible Jenkins Docker CI/CD RESTful APIs Infrastructure-as-a-Service Log Analysis Agile Source Control Management
3+ yrs exp

Senior Staff Site Reliability Engineer

Cisco

Remote (Irvine, CA) +1 19 days ago $192,400$275,800
SRE Linux Administration AWS GCP Azure Python Go Distributed Systems Splunk SPL Indexer Clustering Search Head Clusters KVStore Monitoring Alerting Observability Root Cause Analysis
10+ yrs exp Remote

Senior Site Reliability Engineer

The Federal Reserve

Boston, MA 19 days ago $140,000$210,900
AWS EKS Terraform Python Java Go Docker Ansible CI/CD IaC Linux Shell Scripting Prometheus Grafana CloudWatch OpenSearch Dynatrace Consul Vault S3 RDS Aurora Route 53 ELB ECR

Senior Site Reliability Engineer

Autodesk

Remote (ID) +1 32 days ago $117,000$209,330
Site Reliability Engineering AWS Kubernetes Python Go Java Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace Bash PowerShell FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp Remote

Senior Site Reliability Engineer

Oracle

Pleasanton, CA +2 83 days ago $81,100$187,000
Terraform Chef Ansible Python Java Bash Kubernetes Helm Jenkins Grafana Prometheus CI/CD OCI DevOps
3+ yrs exp