Engineering Manager, Reliability Engineering Flywheel, EDA Infrastructure

Nvidia

Confirmed live today High trust
Remote

Quick summary

Work type
Remote
Location
Redmond, WATXWACAMA
Salary
$224,000–$356,500 / yr
Posted
2 days ago
Freshness
Confirmed live today

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $198k
This role $290k
$136k most similar roles pay here $380k

This role pays more than 92% of similar roles. Most pay $163,416–$232,850 — the shaded band above. At the midpoint, this role pays about $290k versus about $198k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 879 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 859 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Engineering Manager, Reliability Engineering Flywheel, EDA Infrastructure

Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure joins the EDA Infrastructure organization to lead a team responsible for operational processes and platforms across incident management, maintenance, on-call, issue management, and customer-serving readiness. This role involves owning the roadmap and delivery of tools used by teams while partnering with infrastructure and service owners to improve reliability and reduce manual work. The manager will set technical direction, prioritize work, hire and develop engineers, and provide leadership during major incidents. Key responsibilities include establishing consistent practices for production readiness and integrating automation and AI or LLMs to improve triage, knowledge retrieval, and incident analysis. The role focuses on the specific domain of EDA infrastructure, supporting systems that facilitate chip development while managing complex dependencies and demanding availability requirements within a large-scale compute environment.

What does a Engineering Manager earn in California?

Median $290250 from 57 postings across 17 companies.

See salary data

What you'll do

  • Lead a team and own the roadmap for operational processes and platforms from requirements through adoption.
  • Set technical direction and prioritize work across engineering and operational disciplines.
  • Partner with cross-functional teams to establish consistent practices for incident response, maintenance, and on-call management.
  • Develop tools and automation to reduce manual work and improve service reliability.
  • Hire, develop, and manage engineers and technical leads while ensuring clear ownership and accountability.
  • Provide technical leadership and communication during major incidents and high-pressure situations.
  • Integrate AI and LLMs to improve triage, knowledge retrieval, and incident analysis.

What we're looking for

  • BS degree or equivalent experience with software engineering or related experience.
  • 5+ years of engineering leadership managing teams or complex technical programs.
  • Knowledge of operational processes and supporting platforms, including roadmap, delivery, adoption, and improvement.
  • Strong technical judgment in software architecture, platform integration, and engineering tradeoffs.
  • Clear communication with engineers, cross-functional partners, and executive stakeholders.
  • Proven record of developing engineers, growing teams, and delivering results under pressure.
  • Experience with established readiness standards for service ownership, support coverage, and reliability (preferred).
  • Experience building platforms for incident management, maintenance, customer experience, or using AI/LLMs to improve triage (preferred).

More like this

Similar roles

Reliability Engineering Manager

Anduril Industries

Costa Mesa, CA 17 days ago $166,000–$253,000
FMEA FMECA Fault Tree Analysis Weibull Analysis HALT HASS Predictive Maintenance MIL-STD-810 MIL-HDBK-217 MIL-STD-461 MIL-STD-516C MIL-STD-1629 RTCA/DO-330 DO-254 DO-178 MIL-HDBK-472 ARP-4754 ARP-4761 Relyence Teamcenter Confluence JIRA
7+ yrs exp

Senior Engineering Manager, Site Reliability

Upstart

Remote (Canada) 68 days ago $195,300–$270,400
Site Reliability Engineering Distributed Systems Cloud Infrastructure Datadog Grafana Prometheus OpenTelemetry Kubernetes AWS Incident Management Observability Service Level Objectives Postmortem Cloud Native Architecture
7+ yrs exp Remote

Engineering Manager, Cloud Network Reliability

Apple Inc

Sunnyvale, CA 140 days ago $237,600–$356,400
SRE SDN Distributed Systems Microservices RESTful APIs Observability Metrics Logging Tracing Networking Protocols Routing Traffic Management Cloud-Native Platforms Fault-Tolerant Systems
10+ yrs exp

Engineering Manager, Cloud Network Reliability

Apple Inc

Seattle, WA 49 days ago $225,600–$338,400
SRE SDN Distributed Systems Microservices RESTful APIs Observability Metrics Logging Tracing Routing Traffic Management Cloud-Native Platforms Fault-Tolerant Systems Infrastructure Engineering
10+ yrs exp

Manager, Site Reliability Engineering

Okta Inc

San Francisco, CA 56 days ago $204,000–$306,000
AWS Kubernetes Terraform CI/CD Grafana Splunk APM DevOps SaaS Cloud-native Architecture Containerization Edge Networking SDLC
3+ yrs exp Hybrid