Senior Solutions Architect, Cloud Partner Operations

Nvidia

Confirmed live 2 days ago High trust
Remote

Quick summary

Work type
Remote
Location
CA, Canada
Salary
$224,000–$356,500 / yr
Posted
29 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $212k
This role $290k
$147k most similar roles pay here $379k

This role pays more than 91% of similar roles. Most pay $180,100–$244,050 — the shaded band above. At the midpoint, this role pays about $290k versus about $212k for comparable roles.

Based on 239 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Senior Solutions Architect, Cloud Partner Operations

As a Senior Solutions Architect, NVIDIA Cloud Partner Operations, you will join the team to elevate Day 2 operations across the NVIDIA Cloud Partner ecosystem. You will work directly with partner engineers to solve complex problems regarding performance, stability, and economics for AI clouds at scale. Your daily responsibilities include prototyping solutions under representative loads, creating reference architectures, and developing automated workflows to improve reliability and infrastructure maturity. The role requires expertise in large-scale GPU, HPC, or cloud infrastructure, including technologies such as DCGM, BMC/Redfish, InfiniBand, NCCL, Lustre, and Kubernetes. You will utilize tools like Prometheus, Grafana, Terraform, Ansible, and Argo CD while leveraging Python and Bash for automation. This role focuses on the technical challenge of making new technology Day 2 ready by improving unit economics and operational efficiency in live environments.

What does a Solutions Architect earn in California?

Median $235750 from 45 postings across 8 companies.

See salary data

What you'll do

  • Solve complex Day 2 operations problems at scale to improve performance, stability, and economics for cloud partners.
  • Work with partner engineers to identify root causes, prototype solutions, and validate them under representative loads.
  • Prepare operating models for new NVIDIA platforms and services before they are deployed in live environments.
  • Identify and close gaps in partner infrastructure regarding telemetry, security, incident response, and tooling.
  • Convert validated technical solutions into reference architectures, automated workflows, and standard operating procedures for the ecosystem.
  • Analyze metrics like incident frequency and cost per token to identify where clouds are losing performance or margin.
  • Identify patterns across different partners to provide feedback to NVIDIA product and engineering teams for systemic fixes.

What we're looking for

  • BS, MS, or PhD in Computer Science, Engineering, Physics, Mathematics, or a related field, or equivalent experience.
  • 12+ years of experience in production infrastructure, cloud engineering, solutions architecture, SRE, HPC, or similar technical roles.
  • 5+ years of exceptional specialist-level work in large-scale GPU or AI infrastructure (alternative to the 12-year requirement).
  • Experience building, operating, or improving distributed infrastructure under real production load.
  • Deep expertise in at least one part of the Day 2 stack with hands-on experience in large-scale GPU, HPC, or cloud infrastructure.
  • Proficiency in Linux and scripting languages like Python or Bash for automation and troubleshooting.
  • Experience with Kubernetes, Slurm, Prometheus, Grafana, OpenTelemetry, Terraform, Ansible, or Argo CD.
  • Real-world experience operating a GPU cloud, HPC environment, or large-scale AI platform under customer load (preferred).

More like this

Similar roles

Senior Solutions Architect, Cloud Partner Operations

Nvidia

Remote 21 days ago $224,000$356,500
GPU HPC Kubernetes Slurm Terraform Ansible Argo CD Python Bash Linux Prometheus Grafana OpenTelemetry InfiniBand NCCL Lustre IBM Storage Scale WEKA VAST Data DCGM BMC/Redfish
10+ yrs exp Remote

Senior Solutions Architect

Nvidia

Remote 14 days ago $184,000$287,500
Python PyTorch TensorFlow CUDA NVIDIA Triton Inference Server TensorRT TensorRT-LLM MLOps SLURM AWS Azure GCP Deep Learning LLMs HPC NCCL DCGM UFM
10+ yrs exp Remote

Senior Solutions Architect

Nvidia

Santa Clara, CA 149 days ago $184,000$287,500
Generative AI LLMs GPU HPC NCCL DCGM UFM Mission Control Base Command Manager SLURM Networking Compute Infrastructure Performance Analysis AI Benchmarking Data Center Design Accelerated Computing Deep Learning
8+ yrs exp

Solutions Architect

Nvidia

Remote 14 days ago $184,000$287,500
GPU AI HPC GenAI Networking Data Center Design NCCL DCGM UFM Mission Control Base Command Manager SLURM Performance Testing AI Benchmarking
7+ yrs exp Remote

Solutions Architect

Nvidia

Remote 15 days ago $184,000$287,500
GPU AI HPC GenAI Networking Data Center Design NCCL DCGM UFM Mission Control Base Command Manager SLURM Performance Testing AI Benchmarking
7+ yrs exp Remote

Senior Solutions Architect, Cloud Partners - Telco

Nvidia

Remote (Santa Clara, CA) +4 57 days ago $152,000$241,500
Python PyTorch TensorFlow CUDA TensorRT TensorRT-LLM NVIDIA NeMo Triton Inference Server MLOps SLURM AWS Azure GCP Deep Learning LLMs HPC Cloud Architecture
6+ yrs exp Remote