Senior Software Engineer, AI Frameworks

Nvidia

Confirmed live today High trust
Remote

Quick summary

Work type
Remote
Location
Santa Clara, CA
Employment
Full-time
Posted
11 days ago
Freshness
Confirmed live today

Market check

Salary context

How this pay compares to similar roles

Similar $174k
$108k most similar roles pay here $231k

This listing doesn't post a salary. Most similar roles pay $149,500–$197,750.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 1048 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 834 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Software Engineer, AI Frameworks

Senior Software Engineer, AI Frameworks will drive the integration of the NVIDIA Grove project within Dynamo and across various leading open-source AI frameworks. The role involves developing production-grade software to enable the adoption, scaling, and operation of Grove capabilities across environments including llm-d, Ray, PyTorch, and other emerging ecosystem projects. Key responsibilities include building adapters, plugins, and runtime components while partnering with framework owners to upstream changes and develop reference workflows. The candidate will optimize performance for distributed training and inference in multi-node and multi-GPU environments while improving observability for Kubernetes-based deployments. Required skills include proficiency in Go, C++, or Python, along with experience in distributed systems, containerization, and GPU performance profiling. This role addresses the technical challenge of ensuring seamless integration and reliability across complex AI infrastructure and training stacks.

What does a Software Engineer earn in California?

Median $214000 from 763 postings across 66 companies.

See salary data

What you'll do

  • Design and implement end-to-end integrations of Grove with open-source AI frameworks like PyTorch, Ray, and Dynamo.
  • Build and maintain adapters, plugins, and runtime components to enable Grove features across training and inference stacks.
  • Partner with framework owners to upstream changes and ensure long-term maintainability of integrated systems.
  • Develop reference workflows, sample applications, and best-practice guides to accelerate user adoption.
  • Optimize performance, scalability, and reliability for distributed training and inference in multi-node and multi-GPU environments.
  • Improve observability and operational readiness for Kubernetes-based deployments through metrics, logging, and tracing tools.
  • Diagnose complex issues involving containers, networking, scheduling, and CUDA/GPU utilization.
  • Define APIs and contracts to ensure compatibility across various versions of frameworks and dependencies.

What we're looking for

  • BS/MS/PhD in Computer Science, Electrical Engineering, or a related field (or equivalent experience).
  • 5+ years of proven experience in a related field.
  • Hands-on experience integrating with at least one major AI framework or runtime like PyTorch, Ray, or Triton Inference Server.
  • Solid understanding of AI workloads including model development, training vs. inference tradeoffs, and performance considerations.
  • Experience with distributed systems concepts such as RPC, scheduling, fault tolerance, and resource management.
  • Practical experience in Kubernetes, including deploying services, Helm/Kustomize, and debugging clusters.
  • Strong software engineering experience in Go, C++, and/or Python to ship reliable systems.
  • Open-source contributions to Dynamo, PyTorch, Ray, or the Kubernetes ecosystem (preferred).

More like this

Similar roles

Senior Software Engineer, AI Inference Systems

Nvidia

Santa Clara, CA 156 days ago
vLLM SGLang CUDA Python C++ PyTorch Triton MLIR LLVM XLA Docker Kubernetes Slurm NCCL Nsight Systems CI/CD AWS GCP Azure High-Performance Computing
7+ yrs exp Hybrid

Senior Software Engineer

Apple Inc

Seattle, WA 37 days ago $175,000–$263,300
Java Maven Gradle Bazel SQL Relational Databases ORM Kubernetes Containers Machine Learning LLMs

Senior Software Engineer, AI Frameworks

Microsoft

Remote 15 days ago $119,800–$234,700
C++ Python CUDA Triton vLLM SGLang Kernel Optimization Benchmarking Profiling AI Frameworks Observability
4+ yrs exp Remote

Senior AI Frameworks Engineer

Nvidia

Santa Clara, CA +1 126 days ago $152,000–$241,500
C++ Python CUDA LLVM MLIR JIT AST FFI GPU Programming Conda High-Performance Computing DSL NVIDIA GPU
3+ yrs exp