Senior Deep Learning Software Engineer, Inference
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 65% of similar roles. Most pay $188,987–$251,875 — the shaded band above. At the midpoint, this role pays about $236k versus about $220k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 892 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 870 roles with salary data.
Most-posted roles
At a glance
Senior Deep Learning Software Engineer, Inference will join the team to design, build, and optimize GPU-accelerated software powering sophisticated AI applications. The role focuses on developing high-performance deep learning frameworks, specifically SGLang and vLLM, to facilitate the deployment of large language models. You will perform performance optimization, analysis, and tuning for LLM, Multimodal, and Generative AI models across various NVIDIA accelerators from datacenter GPUs to edge SoCs. Key responsibilities include contributing code to inference libraries like FlashInfer and collaborating with cross-functional teams on innovative solutions. Required skills include expert C/C++ programming, software design, and Python. You will utilize tools such as CUTLASS, OAI Triton, NCCL, and CUDA kernels to optimize model serving pipelines. The work addresses the technical challenge of efficient large-scale model serving and high-performance inference for state-of-the-art generative models.
Skills
What you'll do
What we're looking for
More like this
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia