Senior Product Manager AI Inference Software

Amd

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
San Francisco, CAAustin, TX
Salary
$205,680–$308,520 / yr
Posted
72 days ago
Freshness
Confirmed live yesterday
Closes
Jul 1, 2027

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $203k
This role $257k
$151k most similar roles pay here $325k

This role pays more than 85% of similar roles. Most pay $177,250–$228,412 — the shaded band above. At the midpoint, this role pays about $257k versus about $203k for comparable roles.

Based on 240 similar postings.

Employer

About Amd

AMD (Advanced Micro Devices) is a semiconductor company that develops high-performance processors, graphics cards, and adaptive computing solutions for gaming, data centers, and embedded markets. Industry: Semiconductors

Amd currently has 367 open roles on FindRole.

Listed pay typically runs $166,400–$249,600 across 367 roles with salary data.

Most-posted roles

View all roles at Amd

At a glance

TL;DR · Senior Product Manager AI Inference Software

Senior Product Manager –AI Inference Software joins the team to own the framework-layer inference product strategy for ROCm, focusing on production AI inference on AMD Instinct hardware. The role involves translating customer, ecosystem, and engineering signals into roadmap decisions to enable efficient, reliable, and competitive large-scale model deployment. Key responsibilities include defining software stack capabilities for single-GPU through rack-scale deployments, identifying emerging model architectures, and partnering with engineering teams across inference engines, serving, orchestration, memory systems, and GPU libraries. The candidate will engage with the open-source community and sophisticated AI infrastructure customers to drive technical requirements. Required expertise includes production AI inference systems, including vLLM, SGLang, AMD ATOM, and llm-d, alongside knowledge of memory management and distributed inference. This role addresses the challenge of scaling large models within a complex software ecosystem.

What you'll do

  • Own the product strategy and roadmap for ROCm's inference framework software capabilities.
  • Translate customer, ecosystem, and engineering signals into prioritized technical requirements for production AI workloads.
  • Define how the software stack enables large-scale model deployment across single-GPU and rack-scale environments.
  • Identify emerging model architectures and infrastructure shifts to determine where AMD should differentiate in the software stack.
  • Partner with engineering teams to identify performance bottlenecks and drive execution across organizational boundaries.
  • Serve as an active representative for AMD within the open-source inference community and technical forums.
  • Engage directly with AI infrastructure customers to translate their specific requirements into scalable product decisions.
  • Collaborate with cloud providers and model developers to develop joint integrations and solutions.

What we're looking for

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related technical field.
  • Deep understanding of production AI inference systems including vLLM, SGLang, and AMD ATOM.
  • Experience with serving and orchestration systems like llm-d, memory management, and distributed inference.
  • Experience driving complex technical initiatives across multiple engineering teams or organizations.
  • Experience working with sophisticated AI infrastructure customers, open-source communities, or ecosystem partners.
  • Experience in GPU computing, distributed systems, cloud infrastructure, high-performance computing, or technical product leadership.
  • Ability to communicate technical tradeoffs clearly to engineering leaders, customer stakeholders, and executive audiences.
  • Familiarity with AMD Instinct hardware and the ROCm software ecosystem is a plus.

More like this

Similar roles

Senior Product Manager, AI Platform Inference

Nvidia

Santa Clara, CA 37 days ago $168,000$258,750
vLLM SGLang FlashInfer TensorRT-LLM Triton Dynamo TorchAO NVIDIA GPUs Generative AI Machine Learning GPU Architecture Performance Profiling Software Development Open Source GitHub
6+ yrs exp

Senior Product Manager, AI Inference Performance

Nvidia

Santa Clara, CA 29 days ago $208,000$327,750
AI Inference TensorRT-LLM vLLM SGLang NVIDIA Dynamo Triton Inference Server Quantization Speculative Decoding KV Caching LLM Deep Learning Generative AI Benchmarking Capacity Planning Autoscaling
10+ yrs exp