Principal Software Engineer, Compute Infrastructure

Nvidia

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Salary
$248,000–$391,000 / yr
Posted
12 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $198k
This role $320k
$105k most similar roles pay here $422k

This role pays more than 95% of similar roles. Most pay $174,600–$220,800 — the shaded band above. At the midpoint, this role pays about $320k versus about $198k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · Principal Software Engineer, Compute Infrastructure

As a Principal Software Engineer - Compute Infrastructure, you will join the team to define platform architecture and spearhead the operationalization of internal frontier-class AI inference systems. You will lead initiatives to transform a global enterprise compute platform running thousands of nodes and tens of thousands of VMs and containers using OpenShift and KubeVirt. Your daily work involves building automated remediation pipelines, hardware watchdogs, and telemetry for rack-scale GPU systems while managing capacity planning and complex migrations of legacy workloads into modern Kubernetes orchestration. You will utilize Go, Python, Terraform, and OpenTofu to build self-service architectures and GitOps workflows via ArgoCD. The role focuses on solving infrastructure challenges for large-scale AI inference, addressing hardware supply constraints, and ensuring operational maturity across bare metal, virtualized environments, and multi-cloud deployments involving high-speed backplane networking and advanced storage protocols.

What does a Software Engineer earn?

Median $197500 from 2196 postings across 133 companies.

See salary data

What you'll do

  • Architect and transform the global enterprise compute platform using OpenShift and KubeVirt for thousands of nodes.
  • Build the operational foundation and automated remediation pipelines for internal frontier-class AI inference systems.
  • Develop telemetry and hardware watchdogs for pre-release, rack-scale GPU systems including Blackwell architectures.
  • Perform capacity planning and develop proactive scaling strategies to navigate hardware supply constraints.
  • Design self-service architectures, APIs, and Terraform/OpenTofu providers to drive adoption of standard platforms.
  • Lead the migration of massive legacy workloads into modern Kubernetes orchestration environments.
  • Manage large-scale infrastructure across bare metal, virtualized systems, and multi-cloud environments using a GitOps posture.

What we're looking for

  • Bachelor's degree in Engineering, Computer Science, Mathematics, or a related field, or equivalent experience.
  • 15+ years of proven experience in compute platform engineering, site reliability, or systems architecture with a focus on automation at scale.
  • Deep expertise in Kubernetes architecture and designing/deploying virtualization architectures, specifically KubeVirt and OpenShift.
  • In-depth knowledge of hardware technologies including GPUs and high-speed backplane networking to mitigate large-scale failures.
  • Experience managing global environments across bare metal, virtualized infrastructure, and cloud with a unified GitOps posture.
  • Proficiency in programming languages such as Go and/or Python.
  • Expert-level experience in infrastructure-as-code development using Terraform or Config Management.
  • Strong leadership skills to influence technical direction across highly autonomous engineering teams.

More like this

Similar roles

Principal Software Engineer, Infrastructure

Nvidia

Santa Clara, CA +1 17 days ago $248,000$391,000
Ansible Automation Platform AWX Salt Go Python Java Linux Kubernetes Containers CI/CD Infrastructure as Code Databricks Distributed Systems Hybrid Infrastructure Observability Secrets Management
10+ yrs exp

Principal Software Engineer, AI Infra Compute

Oracle

Austin, TX +1 84 days ago $114,600$234,600
Python Java TypeScript OCI AWS Azure GCP Kafka Docker Linux Bash Perl Ruby RESTful APIs Swagger OpenAPI NLP RoCE Infiniband Agile
6+ yrs exp

Principal Core Infrastructure Engineer

Oracle

Nashville, TN 35 days ago $114,600$234,600
Oracle Cloud Infrastructure Microservices Firmware SmartNIC ILOM NVIDIA AMD Intel Automation Pipelines Observability Hardware-Software Integration
8+ yrs exp

Principal Software Engineer, Core Infrastructure

Oracle

Nashville, TN 27 days ago $114,600$234,600
Distributed Systems System Design Infrastructure as Code (IaC) Data Processing Telemetry Monitoring Alerting Fault Injection rate‑limiting Encryption Cloud Infrastructure Scalability Automation

Principal Software Engineer, Core Infrastructure

Oracle

Santa Clara, CA +1 27 days ago $114,600$234,600
Distributed Systems Infrastructure as Code (IaC) Data Processing Telemetry Monitoring Alerting rate‑limiting Fault Injection Encryption Automation Scalability
3+ yrs exp

Principal Software Engineer, Core Infrastructure

Oracle

Nashville, TN 27 days ago $114,600$234,600
Distributed Systems Infrastructure as Code (IaC) Data Processing Telemetry Monitoring Alerting Fault Injection rate‑limiting Encryption Automation Scalability Data Replication Multi-tenant Environments
3+ yrs exp