High-Performance Storage Architect

Nvidia

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
Santa Clara, CA
Salary
$124,000–$195,500 / yr
Employment
Full-time
Posted
38 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $226k
This role $160k
$106k most similar roles pay here $294k

This role pays less than 91% of similar roles. Most pay $191,143–$261,850 — the shaded band above. At the midpoint, this role pays about $160k versus about $226k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 1371 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 1096 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · High-Performance Storage Architect

The High-Performance Storage Architect, NVIS joins the Infrastructure Specialists team to support the development of large-scale AI Factories. This role involves deploying, managing, and validating high-performance storage infrastructure within Linux-based environments while serving as a domain expert during customer planning calls and implementation phases. The architect is responsible for creating technical documentation, performing knowledge transfers, and providing internal feedback through bug reporting and performance optimization suggestions. Key technical requirements include experience with SDS, NFS, S3, NVMeoF, Lustre, GPFS, and RMDA technologies like RoCEv2 or InfiniBand. Candidates must be proficient in Linux system administration, advanced networking, and scripting using Bash, Python, or Ansible. Additionally, the role requires familiarity with benchmarking tools such as FIO, IO500, IOR, HPL, NCCL tests, and MLPerf storage to solve complex data pipeline challenges for AI training and inference.

What you'll do

  • Deploy, manage, and validate high-performance storage infrastructure within Linux-based AI Factory environments.
  • Serve as the primary domain expert during customer planning calls throughout all implementation phases.
  • Create technical documentation and perform knowledge transfers to support customers effectively.
  • Identify and report bugs while documenting workarounds and suggesting improvements to internal teams.
  • Optimize storage performance for AI training, checkpointing, inference, and large-scale data pipelines.
  • Perform system administration, performance reporting, and network routing tuning and monitoring.
  • Execute storage benchmarking using tools such as FIO, IO500, IOR, HPL, NCCL, and MLPerf.

What we're looking for

  • 5+ years of experience providing support, deployment, and validation services for hardware and software-based storage products.
  • Solid understanding of storage concepts and technologies including SDS, NFS, S3, NVMeoF, Lustre, GPFS, and RMDA (RoCEv2/InfiniBand).
  • Familiarity with storage benchmarking tools such as FIO, IO500, IOR, HPL, NCCL tests, and MLPerf storage.
  • Proficiency in Linux system administration, performance reporting/optimization/logging, and advanced network routing.
  • Scripting proficiency in Bash, Python, Ansible, etc.
  • A four-year degree in Computer Science, Electrical or Computer Engineering, or equivalent experience.
  • Experience with parallel and distributed filesystem products such as Ceph, Weka.io, Vast, or DDN (preferred).
  • Expertise in optimizing storage performance for AI training, checkpointing, inference, or large-scale data pipelines (preferred).

More like this

Similar roles

Senior HPC Storage Engineer

Nvidia

Santa Clara, CA +1 2 days ago $184,000–$287,500
Distributed Storage HPC Python Bash Docker Enroot Ceph Weka.io Vast Lustre GPFS CUDA NCCL MLPerf NVMe SDN PyTorch TensorFlow Linux RHEL
8+ yrs exp

Distinguished Engineer, Storage - AI Cloud

Nvidia

Santa Clara, CA 10 days ago $320,000–$488,750
Lustre GPFS Spectrum Scale WEKA VAST BeeGFS DAOS S3 NVMe-oF NVMesh iSCSI C C++ Rust Go Python Linux Kernel RDMA RoCE InfiniBand SPDK NFS Ceph MinIO RocksDB GitHub
10+ yrs exp

Senior Manager, Storage Production Engineering

Nvidia

Remote 67 days ago $272,000–$431,250
Lustre GPFS Ceph MinIO NetApp Pure Storage NVMe-oF RDMA NFS SMB iSCSI Fibre Channel Terraform Ansible Puppet Prometheus InfluxDB Elastic Stack Kubernetes S3 AWS Azure
10+ yrs exp Remote

Senior Site Reliability Engineer, Storage

Nvidia

Santa Clara, CA 33 days ago $168,000–$270,250
HPC Distributed File Systems Lustre GPFS NetApp Pure Storage S3 MinIO Python Bash Golang AWS Azure GCP Prometheus Grafana Elasticsearch Kibana Splunk Zabbix RDMA InfiniBand RoCE Slurm PBS LSF Docker Kubernetes
8+ yrs exp