DGX Cloud Automation Engineer

Nvidia

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Santa Clara, CA
Salary
$184,000–$287,500 / yr
Posted
33 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $175k
This role $236k
$117k most similar roles pay here $306k

This role pays more than 84% of similar roles. Most pay $135,187–$214,562 — the shaded band above. At the midpoint, this role pays about $236k versus about $175k for comparable roles.

Based on 240 similar postings.

Employer

About Nvidia

Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing

Nvidia currently has 896 open roles on FindRole.

Listed pay typically runs $184,000–$287,500 across 876 roles with salary data.

Most-posted roles

View all roles at Nvidia

At a glance

TL;DR · DGX Cloud Automation Engineer

As a DGX Cloud Automation Engineer on the DGX Cloud Engineering Team, you will play a critical role in building development and release processes for AI Factories software while maintaining high standards for CI/CD and release engineering tools. You will design, build, and implement scalable cloud-based systems for PaaS and IaaS environments, collaborating with developers, QA, and product teams to streamline software release cycles. The role requires expertise in Kubernetes, KubeVirt, Docker, and Infrastructure as Code to support both on-premises and cloud deployments. Technical requirements include proficiency in Golang, building RESTful web services, and managing bare metal infrastructure involving PXE Boot, DHCP, DNS, and OS. You will work within the domain of a cloud platform tailored for AI tasks, specifically focused on transitioning projects from development to deployment using GPU-powered services.

What you'll do

  • Build and maintain development and release processes for AI Factories software.
  • Maintain high standards for CI/CD tools and release engineering processes.
  • Design, build, and implement scalable cloud-based systems for PaaS and IaaS.
  • Develop and improve CI/CD tools for both on-premises and cloud deployments.
  • Streamline software release processes in coordination with developers, QA, and product teams.
  • Support, maintain, and document software functionality for the DGX Cloud platform.

What we're looking for

  • BS or MS in Computer Science or equivalent experience from an accredited institution.
  • 8+ years of experience in DevOps and Deployment.
  • 2+ years of programming experience in Golang.
  • Expertise in Kubernetes (K8s) and KubeVirt.
  • Experience with Infrastructure as Code, Docker, Containers, and building RESTful web services.
  • Experience with Continuous Integration and Continuous Delivery tools for on-prem and cloud deployment.
  • Knowledge of cloud design including virtualization, global infrastructure, distributed systems, and security.
  • Background with Cloud Service Providers such as AWS (Fargate, EC2, IAM, ECR, EKS, Route53).
  • Experience bringing up bare metal in a datacenter (PXE Boot, DHCP, DNS, OS) (preferred).
  • Expertise in virtualization technologies like Firecracker, KVM, OpenStack, Nutanix AHV, and Redhat OpenShift (preferred).

More like this

Similar roles

Principal Software Engineer

Nvidia

Remote 35 days ago $272,000$431,250
Kubernetes Go C Linux CI/CD GitLab Argo Flux Container Orchestration Distributed Systems Cloud Computing GPU DPU Confidential Computing
10+ yrs exp Remote

Senior Cloud Software Engineer

Nvidia

Remote 6 days ago $152,000$241,500
Kubernetes AWS GCP Azure Go Python Rust C++ Java Distributed Systems Cloud-Native Data Management Storage Systems Performance Engineering Observability
5+ yrs exp Remote

Principal Software Engineer

Nvidia

Santa Clara, CA +1 135 days ago $272,000$431,250
Go Python Java Kubernetes Slurm Prometheus OpenTelemetry Grafana Docker AWS GCP Azure CUDA cuDNN Distributed Systems Infrastructure Automation Workflow Orchestration
10+ yrs exp

Principal Software Engineer, Distributed Systems Engineer

Nvidia

Remote (Durham, NC) 78 days ago $272,000$431,250
Kubernetes GPU Go Python Slurm Bright Cluster Manager Distributed Systems Cluster Management Monitoring Data Structures Algorithms Systems Programming Network Telemetry Incident Management
10+ yrs exp Remote

Senior Customer Success Engineer

Nvidia

Remote 32 days ago $200,000$322,000
Golang Python C++ Java Rust Kubernetes Distributed Systems Cloud Infrastructure High-Performance Computing DevOps Networking Storage AI/ML Workloads IaaS PaaS SaaS Scripting
10+ yrs exp Remote