Senior HPC Support Engineer, Ethernet and AI Infrastructure

Nvidia

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Memphis, TN
Salary
$108,000–$172,500 / yr
Employment
Full-time
Posted
100 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $187k
This role $140k
$94k most similar roles pay here $242k

This role pays less than 88% of similar roles. Most pay $151,018–$222,000 — the shaded band above. At the midpoint, this role pays about $140k versus about $187k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 1388 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 1103 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior HPC Support Engineer, Ethernet and AI Infrastructure

Senior HPC Support Engineer - Ethernet and AI Infrastructure joins the NVEX Global Technical Support team to provide comprehensive solutions for sophisticated installations, maintenance, and operations of groundbreaking networking products. This role serves as the primary point of contact for a prestige customer, involving technical troubleshooting, research, and issue resolution via phone, email, or conference calls. The position focuses on NVIDIA Ethernet Switching technologies and End-to-End Solutions like Spectrum-X within Linux multi-distro environments. Key responsibilities include collaborating with Engineering and Marketing teams to improve support processes and documenting standard methodologies. Required expertise includes networking protocols such as TCP, UDP, BGP, and OSPF; Linux system administration; and packet analysis using Wireshark or TCPDUMP. Candidates should possess experience in data centers, distributed systems, and containerization while utilizing AI tools like ChatGPT or Copilot to resolve complex hardware and software issues.

What you'll do

  • Resolve complex technical issues and customer concerns regarding NVIDIA Ethernet Switching technologies and Spectrum-X solutions.
  • Provide on-site technical support, debugging, and issue resolution at the customer location in Memphis.
  • Troubleshoot hardware and software issues across multi-distribution Linux operating systems.
  • Analyze network traffic and protocols using tools like TCPDUMP and Wireshark to resolve connectivity problems.
  • Act as the primary point of contact for customers regarding installation, operation, maintenance, and interoperability.
  • Develop and document standard methodologies for internal support and R&D teams to improve support processes.
  • Provide feedback to engineering and marketing teams regarding product requirements and customer experience.
  • Manage high-level resolution for critical issues while maintaining high levels of customer satisfaction.

What we're looking for

  • 5+ years of experience providing in-depth customer support and debugging for hardware and software products.
  • An academic degree in Networking, Computer Science/Engineering, or Electrical/IT (or equivalent experience).
  • Proficiency in Linux system administration and networking across multiple distributions.
  • Expertise in enterprise networking protocols including TCP, UDP, Ethernet, IP, L2, L3, ARP, STP, LACP, MLAG, IGMP, PIM, BGP, and OSPF.
  • Ability to debug networking protocols using tools such as TCPDUMP and Wireshark.
  • Deep understanding of at least two of the following: data centers, servers, distributed systems, virtualization, deep learning frameworks, or containerization.
  • Experience in large-scale networking and AI infrastructure with overlay technologies (BGP, OSPF, VXLAN, EVPN), RoCE, and QoS concepts (preferred).
  • Relevant certifications such as CCIE, JNCIE-DC/ENT, RHCE, LFCS, or NCP-AII/AIO/AIN (preferred).

More like this

Similar roles

Senior HPC Support Engineer, InfiniBand - NVLink

Nvidia

Westford, MA +4 63 days ago $108,000–$172,500
InfiniBand NVLink GPU Technology Linux Ethernet RDMA RoCEv2 TCPDUMP Wireshark Python Bash KVM ESXi AWS OCI NCCL MPI Slurm Shell Scripting
5+ yrs exp

Senior HPC AI Cluster Engineer

Nvidia

Remote (Santa Clara, CA) 48 days ago $176,000–$276,000
HPC AI GPU CUDA Slurm Kubernetes Python Bash Ansible Jenkins InfiniBand Ethernet RDMA Lustre GPFS Weka.io Linux RedHat CentOS Ubuntu AWS Azure Google Cloud VMware KVM
8+ yrs exp Remote

HPC Infrastructure & Cluster Engineer

General Dynamics

Springfield, VA 34 days ago $119,850–$162,150
Linux Run:AI OpenShift Kubernetes InfiniBand SLURM Python Bash SAN HPC Bare-metal Parallel File Systems Cluster Administration Infrastructure Optimization Network Management Storage Area Network Container Orchestration
5+ yrs exp

Senior System Performance Engineer, Data Center/AI

Qualcomm

San Diego, CA 15 days ago $129,500–$194,300
Python C/C++ SoC ARM AI Machine Learning LLM High-Performance Computing power-management Computer Architecture Memory Hierarchy Cache Systems SMMU GIC Coresight-PMU Distributed Computing Profiling Automation
4+ yrs exp