Senior Manager, Performance Engineering - Kernel and Software Platforms

Nvidia

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Santa Clara, CA
Salary
$272,000–$431,250 / yr
Posted
35 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $195k
This role $352k
$117k most similar roles pay here $465k

This role pays more than 99% of similar roles. Most pay $153,737–$235,750 — the shaded band above. At the midpoint, this role pays about $352k versus about $195k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Manager, Performance Engineering - Kernel and Software Platforms

As a Senior Manager, Performance Engineering – Kernel and Software Platforms, you will lead an engineering team responsible for supervising and optimizing product performance throughout the full hardware lifecycle. You will bridge the gap between design-time predictions and real-world execution by tracking performance from pre-silicon simulation and emulation through post-silicon bring-up to production hardware. Key responsibilities include architecting expectation models, curating workload testlists, and analyzing kernel authoring flows from DSLs like Triton and PyTorch through compiler Intermediate Representations such as MLIR, LLVM IR, and NVVM/PTX. You will coordinate with compiler and architecture teams to align with CUDA release schedules while integrating AI tools like Claude and Codex to automate telemetry analysis and root-cause performance deltas. The role requires expertise in Python automation and deep knowledge of the hardware software pipeline.

What you'll do

  • Supervise and maintain performance tracking for GPU products across the entire lifecycle from pre-silicon simulation to production hardware.
  • Architect theoretical and empirical performance models to set targets and correlate simulation predictions with physical hardware telemetry.
  • Evaluate and benchmark performance translation across kernel authoring flows from high-level DSLs through compiler intermediate representations.
  • Create and maintain stress-test suites and workload testlists to detect regressions and validate hardware release candidates.
  • Coordinate with cross-functional teams to align performance achievements with the official CUDA release schedule.
  • Integrate AI tools and LLM workflows to automate telemetry analysis, root-cause identification, and reporting pipelines.

What we're looking for

  • MS or PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent experience.
  • Experience in systems/software performance engineering, platform benchmarking, or a related area.
  • 5+ years of experience leading or managing technical engineering teams.
  • Demonstrated track record tracking performance through the entire hardware pipeline from design-time simulation to post-silicon production.
  • Proven ability to build expectation models and correlate simulation predictions with physical hardware telemetry.
  • Understanding of modern kernel compilation pipelines, compiler flows (DSL to IR), and how high-level software impacts execution efficiency.
  • Experience developing workload testlists to detect regressions and aligning performance delivery with major software release cycles.
  • Proficiency in Python automation and experience using generative AI models to automate triage and analytical workflows.

More like this

Similar roles

Senior Performance Engineer

Nvidia

Remote (Santa Clara, CA) +2 45 days ago $224,000$356,500
C++ Python CUDA PyTorch JAX XLA GPU Computing Distributed Systems Performance Engineering Benchmarking Profiling Observability High-Performance Computing Data Analysis Automation Workflows
Remote

Senior Performance Engineer

Samsung Semiconductor

San Jose, CA 73 days ago $138,000$206,000
LLM NVIDIA GPU PyTorch vLLM SGLang TensorRT-LLM DeepSpeed Ray Megatron-LM Python C++ Nsight Systems Nsight Compute High-Performance Computing Distributed Systems
5+ yrs exp

Senior Performance Engineer

AES Corporation

Remote (HI) +3 15 days ago $113,000$141,525
Photovoltaic Microgrids SCADA Data Acquisition PowerBI Machine Learning Advanced Analytics Excel Thermal Imaging Drones Grid-forming Frequency/Voltage Regulation
5+ yrs exp Remote

Senior Systems Performance Engineer

Nvidia

Santa Clara, CA 7 days ago $136,000$212,750
CUDA TensorRT Slurm Python vLLM cuBLAS x86 Arm Deep Learning LLM System Architecture Networking Storage
5+ yrs exp

Senior System Performance Engineer

Qualcomm

San Diego, CA 77 days ago $122,500$183,700
Python C/C++ SoC CUDA Vulkan OpenGL DX12 Android Linux Windows AI LLM Power Management System Interconnects SMMU GIC Coresight-PMU Automation
4+ yrs exp