Senior Principal Engineer, AI Networking

Oracle

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$135,200–$306,400 / yr
Posted
122 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $188k
This role $221k
$115k $327k
below market most similar roles pay here above market

This role pays more than 72% of similar roles. Most pay $149,612–$226,350 — the blue band above. At the midpoint, this role pays about $221k versus about $188k for comparable roles.

Based on 240 similar postings.

Employer

About Oracle

Oracle Corporation is a leading multinational technology company specializing in database software, cloud computing, and enterprise software.

Oracle currently has 561 open roles on FindRole.

Listed pay typically runs $102,300–$234,600 across 436 roles with salary data.

Most-posted roles

View all roles at Oracle

At a glance

TL;DR · Senior Principal Engineer, AI Networking

The Senior Principal Engineer - AI Networking joins the team to drive architecture, design, and performance optimization for software components supporting thousands of GPUs and high-bandwidth network fabrics. This role involves architecting high-performance networking software for large-scale AI and HPC environments, specifically designing RDMA-based services to enable low-latency, high-throughput communication across GPU clusters. You will develop congestion management, traffic engineering, and resiliency mechanisms while optimizing end-to-end communication performance across networking and software stacks. Key technologies include C/C++, RDMA, RoCE, InfiniBand, and collective communication libraries like NCCL, MPI, and UCX. You will solve complex problems in distributed systems and networking to support distributed AI training and inference workloads, building foundational infrastructure for next-generation AI workloads while mentoring engineers and defining long-term technology strategy.

What does a Engineer earn in Washington?

Median $179800 from 40 postings across 12 companies.

See salary data

What you'll do

  • Architect and develop high-performance networking software for large-scale AI and HPC environments.
  • Design and implement RDMA-based services to enable low-latency communication across GPU clusters.
  • Drive the evolution of collective communication frameworks and transport layers for distributed AI workloads.
  • Develop congestion management, traffic engineering, and load balancing mechanisms for large-scale RDMA networks.
  • Optimize end-to-end communication performance across networking, GPU, and software stacks.
  • Lead performance analysis, bottleneck identification, and system-wide optimization efforts.
  • Build observability, monitoring, telemetry, and debugging capabilities for large-scale distributed systems.
  • Drive reliability, fault tolerance, and recovery mechanisms for mission-critical AI infrastructure.

What we're looking for

  • Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or related field; advanced degree preferred.
  • 10+ years of software engineering experience building distributed systems, networking software, or infrastructure platforms.
  • Deep expertise in RDMA technologies including RoCE, InfiniBand, or equivalent high-performance networking technologies.
  • Strong experience developing networking software in C/C++.
  • Experience designing and optimizing distributed communication frameworks and transport protocols.
  • Solid understanding of operating systems, networking stacks, memory management, and performance optimization.
  • Experience troubleshooting and optimizing large-scale production systems.
  • Demonstrated technical leadership driving architecture and execution across multiple teams.

More like this

Similar roles

Principal Architect, AI Networking

Nvidia

Remote (Santa Clara, CA) +4 170 days ago $272,000–$431,250
RDMA NVLink GPUDirect InfiniBand RoCE NCCL UCX MPI NVSHMEM C C++ Rust Python CUDA vLLM SGLang TensorRT-LLM Distributed Training Model Parallelism Computer Architecture
10+ yrs exp Remote

Principal Software Engineer, AI Networking

Nvidia

Santa Clara, CA 96 days ago
C/C++ RDMA RoCE InfiniBand DPDK DOCA NCCL CUDA Networking Protocols Distributed Systems Firmware Performance Tuning Congestion Control QoS Telemetry Automation
10+ yrs exp

Senior Manager, AI Network Engineering

Oracle

Nashville, TN 38 days ago $169,800–$355,400
RDMA RoCEv2 NCCL MPI UCX CUDA GPUDirect RDMA Ethernet PCIe DMA SmartNICs DPUs NIC Firmware HPC Distributed Computing Collective Communication Congestion Control Performance Engineering
10+ yrs exp

Senior Software Engineer, AI Networking

Nvidia

Austin, TX +1 136 days ago $184,000–$287,500
C/C++ RoCE InfiniBand RDMA DPDK DOCA NCCL CUDA Networking Protocols Distributed Systems Firmware OS Performance Tuning Congestion Control QoS Telemetry Automation
8+ yrs exp

Senior Software Engineer, AI Networking

Nvidia

Seattle, WA 94 days ago $184,000–$287,500
C/C++ RDMA DPDK DOCA NCCL CUDA RoCE InfiniBand Networking Protocols Distributed Systems Firmware OS Performance Tuning Congestion Control QoS Telemetry Automation
8+ yrs exp

Principal Software Engineer, Networking

Microsoft

Remote 4 days ago $142,800–$274,800
Kubernetes Go Rust C C++ Python Linux BGP eBPF WireGuard VXLAN GENEVE InfiniBand RDMA Cilium CI/CD Azure Containers GitHub GitLab
6+ yrs exp Remote