Engineering Manager, Deep Learning Inference
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 83% of similar roles. Most pay $184,156–$270,000 — the shaded band above. At the midpoint, this role pays about $290k versus about $227k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 896 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.
Most-posted roles
At a glance
As the Engineering Manager, Inference Benchmarking — AI Perf, you will serve as a Technical Lead Manager within the Dynamo organization to advance the AIPerf open-source benchmarking platform. You will lead an engineering team in building core infrastructure for load generation, ZMQ-based microservices, and GPU telemetry using DCGM, PyNVML, Prometheus metrics, and Kubernetes-native deployments. Your role involves ensuring the accuracy of benchmark results for LLM, multimodal, diffusion, and computer vision inference while managing upstream integrations with vLLM, TRT-LLM, and SGLang. You will mentor engineers in a high-velocity open-source environment to solve problems regarding production inference decisions like cost optimization and latency reduction. The role requires expertise in systems engineering, distributed systems, and deep knowledge of LLM inference mechanics including KV caching, speculative decoding, and measurement reproducibility across datacenter, local, and edge use cases.
What does a Engineering Manager earn in California?
Median $290250 from 64 postings across 19 companies.
Skills
What you'll do
What we're looking for
More like this
Nvidia
Nvidia
Apple Inc
Apple Inc
Nvidia
MongoDB