Software Engineering Manager, GPU Communications Libraries

Nvidia

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Santa Clara, CA
Salary
$184,000–$287,500 / yr
Posted
8 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $225k
This role $236k
$159k most similar roles pay here $301k

This role pays more than 62% of similar roles. Most pay $193,935–$255,250 — the shaded band above. At the midpoint, this role pays about $236k versus about $225k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Software Engineering Manager, GPU Communications Libraries

Software Engineering Manager - GPU Communications Libraries will join the GPU Communications Libraries and Networking team to lead, mentor, and grow a library engineering team. This technical leadership role involves managing the NVSHMEM and UCX libraries, where the manager is responsible for project planning, execution, quality, and performance. The individual will participate in feature design, collaborate with internal and external partners to define product roadmaps, and identify improvements in infrastructure and processes. The role requires expertise in C/C++, Linux, and systems software fundamentals like computer architecture and operating system principles. Candidates should possess experience in high-performance networking, RDMA, InfiniBand, and Ethernet. The work focuses on solving communication performance challenges for Deep Learning and HPC applications running at massive scales across thousands of GPUs connected via NVLink, PCIe, and high-speed networks to ensure optimal end-to-end application performance.

What does a Software Engineering Manager earn in California?

Median $281300 from 35 postings across 13 companies.

See salary data

What you'll do

  • Lead, mentor, and grow the library engineering team.
  • Manage the planning and execution of projects for NVSHMEM and UCX libraries.
  • Ensure the quality and performance of communication libraries used in DL and HPC applications.
  • Participate in technical feature design and implementation.
  • Interact with internal and external partners to identify use cases and requirements.
  • Collaborate with engineering and product management teams to define the product roadmap.
  • Identify and implement improvements in team processes, infrastructure, and practices.

What we're looking for

  • Experience in the software industry with a specialization in HPC networking or system software.
  • 4+ years of management experience.
  • BS, MS, or Ph.D. in CS, CE, EE, or a related technical field.
  • Experience in systems software, communication runtimes, or high-performance networking software development.
  • Strong understanding of computer architecture, operating system principles, and hardware-software interactions.
  • Excellent C/C++ programming and debugging skills in Linux.
  • Experience with parallel programming models (MPI, SHMEM) and at least one communication runtime (NCCL, NVSHMEM, UCX).
  • Knowledge of RDMA, high-performance networking technologies (InfiniBand, RoCE), and HPC or ML/DL fundamentals.

More like this

Similar roles

Manager, Software Engineering - Cumulus Linux

Nvidia

Santa Clara, CA 15 days ago $224,000$356,500
Linux C Python Debian Networking gNMI OpenConfig Streaming Telemetry Data Modeling Embedded Software Ethernet Kernel Networking Cloud Native System Monitoring Infrastructure Services
10+ yrs exp

Senior System Software Engineer, GPU Performance

Nvidia

Remote 6 days ago $152,000$241,500
HPC C++ Python CUDA NCCL UCX NVSHMEM MPI Infiniband Ethernet RDMA PyTorch TensorFlow Kubernetes Docker SLURM Ansible Performance Engineering Parallel Programming
Remote

Principal Deep Learning Communication Architect

Nvidia

Remote (Santa Clara, CA) +1 150 days ago $272,000$431,250
NCCL UCX UCC NVSHMEM CUDA InfiniBand RDMA RoCE TensorRT-LLM vLLM SGLang Megatron-Core DeepSpeed JAX XLA PyTorch Distributed Ray HPC High-Performance Computing
10+ yrs exp Remote

Senior Software Engineer, NCCL

Nvidia

Santa Clara, CA 3 days ago $152,000$241,500
C++ C CUDA NVIDIA GPUs PyTorch TensorFlow NCCL UCX MPI OpenSHMEM Linux High-Performance Computing InfiniBand iWARP Parallel Programming Deep Learning Frameworks
5+ yrs exp