Senior Solutions Architect, Cloud Partner Operations

Nvidia

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
Remote
Salary
$224,000–$356,500 / yr
Posted
21 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $211k
This role $290k
$147k most similar roles pay here $379k

This role pays more than 91% of similar roles. Most pay $178,900–$244,050 — the shaded band above. At the midpoint, this role pays about $290k versus about $211k for comparable roles.

Based on 239 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Solutions Architect, Cloud Partner Operations

As a Senior Solutions Architect, NVIDIA Cloud Partner Operations, you will join the team to elevate Day 2 operations across the NVIDIA Cloud Partner ecosystem. You will work alongside partner engineers to solve complex problems regarding performance, stability, and economics for AI clouds at scale. Your daily responsibilities include prototyping solutions under representative loads, creating reference architectures, and developing automated workflows to improve reliability and infrastructure maturity. The role requires expertise in large-scale GPU, HPC, or cloud infrastructure. You will utilize technologies such as DCGM, BMC/Redfish, InfiniBand, NCCL, UFM, Lustre, WEKA, VAST Data, Kubernetes, Slurm, Prometheus, Grafana, OpenTelemetry, Terraform, Ansible, and Argo CD. Proficiency in Linux, Python, and Bash is essential for automating diagnosis and remediation while addressing technical challenges related to multi-tenancy, GPU scheduling, and high-performance storage systems.

What does a Solutions Architect earn in Remote?

Median $231100 from 71 postings across 13 companies.

See salary data

What you'll do

  • Solve complex Day 2 operations problems at scale for the NVIDIA Cloud Partner ecosystem.
  • Collaborate with partner engineers to identify root causes and prototype solutions under representative loads.
  • Prepare new NVIDIA platforms and services for production by establishing stable operating models.
  • Improve cloud reliability, performance, and economics using metrics like incident frequency and cost per token.
  • Identify and close gaps in partner infrastructure across telemetry, security, and incident response.
  • Convert validated technical solutions into reference architectures, automation scripts, and standard operating procedures.
  • Analyze patterns across multiple partners to provide feedback to NVIDIA product and engineering teams.

What we're looking for

  • BS, MS, or PhD in Computer Science, Electrical or Computer Engineering, Physics, Mathematics, or a related field, or equivalent experience.
  • 12+ years of experience in production infrastructure, cloud engineering, solutions architecture, site reliability engineering, HPC, or a similar technical role.
  • 5+ years of exceptional specialist-level work in large-scale GPU or AI infrastructure (alternative to the 12-year requirement).
  • Experience building, operating, or improving distributed infrastructure under real production load.
  • Deep expertise in at least one part of the Day 2 stack involving large-scale GPU, HPC, or cloud infrastructure.
  • Proficiency with Linux and sufficient Python or Bash experience to automate measurement, diagnosis, validation, or remediation.
  • Experience with Kubernetes, Slurm, Prometheus, Grafana, OpenTelemetry, Terraform, Ansible, or Argo CD.
  • Experience operating a GPU cloud, HPC environment, or large-scale AI platform under customer load (preferred).

More like this

Similar roles

Senior Solutions Architect, Cloud Partner Operations

Nvidia

Remote (CA, Canada) 29 days ago $224,000$356,500
GPU HPC Kubernetes Slurm Terraform Ansible Argo CD Python Bash Linux Prometheus Grafana OpenTelemetry InfiniBand NCCL Lustre IBM Storage Scale WEKA VAST Data DCGM
10+ yrs exp Remote

Senior Solutions Architect

Nvidia

Santa Clara, CA 149 days ago $184,000$287,500
Generative AI LLMs GPU HPC NCCL DCGM UFM Mission Control Base Command Manager SLURM Networking Compute Infrastructure Performance Analysis AI Benchmarking Data Center Design Accelerated Computing Deep Learning
8+ yrs exp

Senior Solutions Architect

Nvidia

Remote 14 days ago $184,000$287,500
Python PyTorch TensorFlow CUDA NVIDIA Triton Inference Server TensorRT TensorRT-LLM MLOps SLURM AWS Azure GCP Deep Learning LLMs HPC NCCL DCGM UFM
10+ yrs exp Remote

Solutions Architect

Nvidia

Remote 14 days ago $184,000$287,500
GPU AI HPC GenAI Networking Data Center Design NCCL DCGM UFM Mission Control Base Command Manager SLURM Performance Testing AI Benchmarking
7+ yrs exp Remote

Solutions Architect

Nvidia

Remote 15 days ago $184,000$287,500
GPU AI HPC GenAI Networking Data Center Design NCCL DCGM UFM Mission Control Base Command Manager SLURM Performance Testing AI Benchmarking
7+ yrs exp Remote

Senior Solutions Architect, Cloud Partners - Telco

Nvidia

Remote (Santa Clara, CA) +4 57 days ago $152,000$241,500
Python PyTorch TensorFlow CUDA TensorRT TensorRT-LLM NVIDIA NeMo Triton Inference Server MLOps SLURM AWS Azure GCP Deep Learning LLMs HPC Cloud Architecture
6+ yrs exp Remote