Senior Product Manager, AI Platform Inference

Nvidia

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Santa Clara, CA
Salary
$168,000–$258,750 / yr
Posted
37 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $208k
This role $213k
$157k most similar roles pay here $270k

This role pays more than 57% of similar roles. Most pay $185,375–$231,187 — the shaded band above. At the midpoint, this role pays about $213k versus about $208k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Product Manager, AI Platform Inference

As a Senior Product Manager - AI Platform Inference, you will join the Product Management team to build tools, SDKs, and libraries that enable developers to successfully deploy inference models on NVIDIA GPUs. You will be responsible for creating products to improve deployment quality, developing product strategies, establishing roadmaps, and crafting go-to-market plans. Your daily work involves collaborating with internal and external developers to build roadmaps specifically for model optimization software while aligning with leadership to drive company strategy. The role requires expertise in inference deployment and optimization software such as vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, and TorchAO. You must possess deep knowledge of Generative AI, machine learning performance optimization, GPU architecture, hardware/software co-design, and performance profiling to solve complex challenges in the rapidly evolving inference landscape.

What does a Product Manager earn in California?

Median $223300 from 163 postings across 44 companies.

See salary data

What you'll do

  • Build tools, SDKs, and libraries to enable developer inference deployments on NVIDIA GPUs.
  • Develop product strategies, roadmaps, and go-to-market plans for the inference platform.
  • Collaborate with internal and external developers to build roadmaps for model optimization software.
  • Identify key improvements in the inference landscape through direct engagement with the developer community.
  • Align with and drive company strategy in coordination with NVIDIA leadership.
  • Create products that help developers optimize performance, safety, and cost for GenAI deployments.

What we're looking for

  • BS or MS degree in Computer Science, Computer Engineering, or a related field.
  • 6+ years of technical product management experience at a technology company.
  • Experience with inference deployment and optimization software such as vLLM, SGLang, TensorRT-LLM, or Triton.
  • Demonstrable knowledge of GenAI or machine learning concepts, specifically regarding performance optimization.
  • Proficiency in software development and delivery processes.
  • Strong communication and interpersonal skills.
  • Knowledge of GPU architecture, hardware/software co-design, and performance profiling.
  • Experience with open-source, GitHub-first developer products involving deep customer interactions.

More like this

Similar roles

Senior Product Manager, AI Inference Performance

Nvidia

Santa Clara, CA 29 days ago $208,000$327,750
AI Inference TensorRT-LLM vLLM SGLang NVIDIA Dynamo Triton Inference Server Quantization Speculative Decoding KV Caching LLM Deep Learning Generative AI Benchmarking Capacity Planning Autoscaling
10+ yrs exp

Senior Deep Learning Software Engineer, Inference

Nvidia

Remote (CA) +4 15 days ago $152,000$241,500
CUDA C++ Python PyTorch vLLM SGLang Triton CUTLASS NCCL NVSHMEM Deep Learning LLM Generative AI GPU Programming Performance Optimization Profiling Agile
5+ yrs exp Remote

Engineering Manager, Deep Learning Inference

Nvidia

Santa Clara, CA +4 37 days ago $184,000$287,500
CUDA Triton CUTLASS C++ Python vLLM SGLang FlashInfer TensorRT-LLM PyTorch NCCL NVSHMEM NIXL Multi-GPU Communication Deep Learning Inference Agile
6+ yrs exp Hybrid

Engineering Manager, Deep Learning Inference

Nvidia

Santa Clara, CA 45 days ago $224,000$356,500
CUDA Triton CUTLASS C++ Python vLLM SGLang FlashInfer TensorRT-LLM PyTorch NCCL NVSHMEM NIXL Multi-GPU Distributed Inference Agile
6+ yrs exp Hybrid

Senior Software Engineer, AI Inference Performance

Nvidia

Santa Clara, CA 16 days ago $184,000$287,500
CUDA Python C++ Rust TensorRT-LLM vLLM SGLang Triton PyTorch NVIDIA Nsight Systems CUTLASS NCCL Quantization Distributed Systems GPU Architecture Model Parallelism Speculative Decoding
6+ yrs exp

Technical Marketing Engineer, AI Platform Software

Nvidia

Santa Clara, CA 43 days ago $136,000$212,750
PyTorch TensorRT-LLM Megatron JAX vLLM SGLang Python C/C++ cuDNN NCCL Deep Learning Machine Learning Multi-GPU Quantization Benchmarking Software Development
4+ yrs exp Hybrid