Senior AI Compute Engineer

Nvidia

Confirmed live yesterday Trusted
Remote

Quick summary

Work type
Remote
Location
Remote
Salary
$148,000–$235,750 / yr
Posted
151 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $207k
This role $192k
$135k most similar roles pay here $269k

This role pays less than 65% of similar roles. Most pay $169,887–$245,112 — the shaded band above. At the midpoint, this role pays about $192k versus about $207k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior AI Compute Engineer

The Senior AI Compute Engineer - NVIS joins the Infrastructure Specialists team to manage and validate large-scale AI Compute and HPC infrastructure in Linux environments. This role involves interacting with customers and internal teams to analyze, define, and implement complex projects involving networking, system design, automation, and validation. The engineer will serve as a domain expert during planning calls, provide technical documentation, and perform knowledge transfers for sophisticated systems. Key responsibilities include troubleshooting hardware and software products while providing feedback to internal teams. Required skills include Linux system administration, kernel management, network routing, and scripting in Bash, Python, or Ansible. Technical requirements include experience with schedulers like SLURM or LSF, Kubernetes, and benchmarking tools such as MLPerf. Preferred expertise includes InfiniBand, GPU hardware, MPI, and storage technologies like Lustre or GPFS.

What you'll do

  • Deploy, manage, and validate AI Compute and HPC infrastructure in Linux-based environments.
  • Act as the primary domain expert for customers during planning calls through project implementation.
  • Create handover documentation and perform knowledge transfers to support customers during system rollouts.
  • Provide technical feedback to internal teams by opening bugs and documenting workarounds.
  • Analyze, define, and implement large-scale AI Compute projects involving networking and system design.
  • Perform performance reporting, optimization, and logging for high-performance computing systems.
  • Automate infrastructure tasks using scripting languages like Bash, Python, or Ansible.

What we're looking for

  • 8+ years of experience providing in-depth support and deployment services for hardware and software products.
  • A four-year degree in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience.
  • Proficiency in Linux system administration including process management, kernel management, and network routing.
  • Proficiency in scripting languages such as Bash, Python, and Ansible.
  • Experience with cluster management and provisioning technologies for bare-metal servers.
  • Experience with schedulers such as SLURM, LSF, or UGE.
  • Experience with benchmarking tools like HPL, NCCL tests, MLPerf, and Kubernetes.
  • Preferred experience with InfiniBand, GPU hardware/software, MPI, and storage technologies like Lustre or GPFS.

More like this

Similar roles

Senior AI Compute Engineer

Nvidia

Remote 66 days ago $148,000$235,750
Linux Python Bash Ansible Kubernetes InfiniBand MPI SLURM LSF UGE HPC NCCL MLPerf HPL Lustre GPFS Networking GPU System Administration Automation
8+ yrs exp Remote

Senior Solutions Architect, AI Compute

Nvidia

Remote 21 days ago $184,000$287,500
Linux Python Bash Ansible Kubernetes SLURM LSF UGE InfiniBand MPI HPL NCCL MLPerf Lustre GPFS GPU HPC Networking System Administration Automation
8+ yrs exp Remote

Senior HPC AI Cluster Engineer

Nvidia

Remote 22 days ago $176,000$276,000
HPC AI GPU CUDA Slurm Kubernetes Python Bash Ansible Jenkins InfiniBand Ethernet RDMA Lustre GPFS Weka.io Linux RedHat CentOS Ubuntu AWS Azure Google Cloud VMware KVM
8+ yrs exp Remote

Senior Manager, Validation and HPC

Nvidia

Remote (TX) +4 14 days ago $216,000$345,000
HPC InfiniBand Ethernet GPU MPI NCCL OpenMP Slurm Salt xCAT Python Bash Perl Linux Unix Ansible Puppet Lustre GPFS
10+ yrs exp Remote