Senior Product Manager, Inference Platform

Nvidia

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Santa Clara, CA
Salary
$208,000–$327,750 / yr
Employment
Full-time
Posted
5 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $213k
This role $268k
$151k most similar roles pay here $347k

This role pays more than 89% of similar roles. Most pay $182,621–$243,600 — the shaded band above. At the midpoint, this role pays about $268k versus about $213k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 1150 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 912 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Product Manager, Inference Platform

As a Senior Product Manager, Inference Platform, you will join the team building foundations for serving AI workloads at scale for internal employees and DSX Partners. You will define product vision and strategy for inference platform capabilities, including APIs, capacity management, cost management, performance optimization, and model serving infrastructure. Your daily work involves translating user needs into prioritized roadmaps, managing the end-to-end product lifecycle, and navigating ambiguous problem spaces to deliver reliable, efficient large-scale model serving. You will utilize expertise in AI/ML systems and inference serving frameworks like TensorRT-LLM, vLLM, or Triton Inference Server. The role requires deep knowledge of model formats, token economics, and advanced techniques such as speculative decoding and KV cache management. You will solve complex problems regarding hardware efficiency, developer experience, and infrastructure constraints for both open source and proprietary models.

What does a Product Manager earn in California?

Median $221750 from 204 postings across 45 companies.

See salary data

What you'll do

  • Define the product vision and strategy for inference platform capabilities including APIs, capacity management, and cost optimization.
  • Translate user needs and technical infrastructure constraints into clear requirements and prioritized roadmaps.
  • Manage the end-to-end product lifecycle for inference-related products and platform investments from concept through launch.
  • Analyze the inference ecosystem to understand model formats, serving frameworks, and performance trade-offs at scale.
  • Resolve ambiguity by framing complex problems, identifying key information, and proposing clear paths forward.
  • Advocate for the user by synthesizing qualitative and quantitative signals into actionable product decisions.
  • Track and synthesize developments in open source models, serving frameworks, and competitive dynamics.

What we're looking for

  • 12+ years of experience delivering complex technical products.
  • Bachelor's degree or higher, or equivalent experience.
  • Strong written and verbal communication skills to create clear product documents, specs, and strategies.
  • Ability to operate in ambiguous, fast-paced environments and make progress without a full playbook.
  • Analytical professional capable of breaking down complex problems and driving toward decisions.
  • Deep familiarity with AI/ML systems and inference serving frameworks like TensorRT-LLM, vLLM, or Triton Inference Server.
  • Experience with inference APIs including design, versioning, performance, hardware efficiency, and developer experience.
  • Familiarity with sophisticated inference techniques (e.g., speculative decoding) and large-scale cloud infrastructure (preferred).

More like this

Similar roles

Senior Product Manager, AI Inference Performance

Nvidia

Santa Clara, CA 51 days ago $208,000–$327,750
AI Inference TensorRT-LLM vLLM SGLang NVIDIA Dynamo Triton Inference Server Quantization Speculative Decoding KV Caching LLM Deep Learning Generative AI Benchmarking Capacity Planning Autoscaling
10+ yrs exp

Senior Product Manager, AI Platform Inference

Nvidia

Santa Clara, CA 59 days ago $168,000–$258,750
vLLM SGLang FlashInfer TensorRT-LLM Triton Dynamo TorchAO NVIDIA GPUs Generative AI Machine Learning GPU Architecture Performance Profiling Software Development Open Source GitHub
6+ yrs exp

Senior Product Manager AI Inference Software

Amd

San Francisco, CA +1 93 days ago $205,680–$308,520
ROCm vLLM SGLang AMD ATOM llm-d GPU Computing Distributed Systems Cloud Infrastructure High-Performance Computing Memory Management open-source Inference Engines Technical Writing
Hybrid

Senior Solutions Architect, AI Inference

Nvidia

Remote (Santa Clara, CA) 37 days ago $184,000–$287,500
TensorRT-LLM vLLM SGLang NVIDIA Dynamo Triton Inference Server Kubernetes Quantization Speculative Decoding KV Cache Management DevOps SFT DPO GRPO RLVR Kernel Optimization Distributed Systems
6+ yrs exp Remote

Senior Product Manager, AI & Analytics

Motorola Solutions

Remote (Waltham, MA) 51 days ago $150,000–$180,000
Artificial Intelligence Computer Vision LLMs Vision-Language Models (VLMs) Agentic AI RAG Retrieval-Augmented Generation Video Management Systems (VMS) IoT
5+ yrs exp Remote