Senior Manager, AI Network Engineering

Oracle

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Nashville, TN
Salary
$169,800–$355,400 / yr
Posted
12 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $196k
This role $263k
$121k most similar roles pay here $380k

This role pays more than 85% of similar roles. Most pay $156,650–$235,750 — the shaded band above. At the midpoint, this role pays about $263k versus about $196k for comparable roles.

Based on 240 similar postings.

Employer

About Oracle

Oracle Corporation is a leading multinational technology company specializing in database software, cloud computing, and enterprise software.

Oracle currently has 666 open roles on FindRole.

Listed pay typically runs $102,300–$223,999 across 619 roles with salary data.

Most-posted roles

View all roles at Oracle

At a glance

TL;DR · Senior Manager, AI Network Engineering

As the Senior Manager - AI Network Engineering, you will lead a team responsible for designing, developing, and troubleshooting software programs for databases, applications, tools, and networks. You will manage the end-to-end NPI lifecycle for high-performance NICs supporting GPU and AI infrastructure while leading the engineering organization for the Collective Communication Library. Your daily work involves driving optimization for RDMA, GPU networking, and topology-aware algorithms across clusters and superclusters. You will oversee hardware/software co-design involving firmware, drivers, and network fabrics to eliminate bottlenecks in large-scale distributed training systems. The role requires expertise in Ethernet, RoCEv2, PCIe, DMA, NCCL, MPI, UCX, GPUDirect RDMA, CUDA, and SmartNICs/DPUs. You will manage technical roadmaps and partnerships to ensure production readiness for complex high-performance networking environments involving thousands of accelerators.

What you'll do

  • Manage the end-to-end NPI lifecycle for high-performance NICs supporting OCI GPU and AI infrastructure.
  • Lead the engineering organization responsible for the architecture, strategy, and roadmap of the Collective Communication Library (CCL).
  • Optimize collective communication, RDMA, GPU networking, and topology-aware algorithms across clusters and superclusters.
  • Develop benchmarking, modeling, and performance-analysis capabilities to identify and eliminate hardware and software bottlenecks.
  • Drive hardware/software co-design across NICs, firmware, drivers, network fabrics, and runtimes.
  • Establish automated qualification, regression, performance, reliability, and fault-testing processes for cloud-scale production.
  • Manage external technology partnerships and communicate technical strategy and risks to senior leadership.

What we're looking for

  • BS or MS in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • 15+ years of experience in systems, networking, distributed computing, HPC, GPU infrastructure, or related areas.
  • Significant experience leading engineering teams responsible for complex hardware/software systems.
  • Deep expertise in high-performance networking including Ethernet, RDMA, RoCE, PCIe, DMA, and modern NIC architectures.
  • Strong knowledge of GPU systems, distributed GPU communication, collective communication algorithms, and large-scale training or HPC systems.
  • Experience taking new hardware technologies from early engineering stages through qualification and large-scale production deployment.
  • Strong systems performance-analysis skills across hardware, firmware, operating systems, networking, runtimes, and applications.
  • Excellent technical and executive communication skills.

More like this

Similar roles

Senior Manager, AI Engineering

Global Payments (TSYS)

Alpharetta, GA 33 days ago
Machine Learning Natural Language Processing AWS GCP Snowflake Apache Spark Apache Kafka CI/CD Data Governance Real-time Data Processing

Senior Engineering Manager, Network Connectivity

Cloudflare, Inc

Portugal 4 days ago $89,000$122,000
Layer 3 Layer 4 Distributed Systems Network Performance Monitoring Incident Response On-call Rotation Automation AI Tools Data Planes Control-plane Services Networking Software Infrastructure
5+ yrs exp Hybrid

Senior Manager, Network Operations

Marqeta

Remote (Ontario, Canada) +1 28 days ago $177,100$221,400
AWS Terraform Cisco Palo Alto Networks BGP OSPF Infrastructure-as-Code SD-WAN Transit Gateway Direct Connect CloudFront Route53 Meraki Network Automation Monitoring
7+ yrs exp Remote

Senior Network Engineer

Coinbase

Remote 10 days ago $186,065$218,900
BGP OSPF Multicast VRFs PIM IGMP Arista MLAG Cisco vPC Fortinet Palo Alto IPsec VPN SSL VPN SNMP gRPC Python Ansible Linux Network Automation
8+ yrs exp Remote

Senior Network Engineer

Leidos

Columbia, MD 159 days ago $116,350$210,325
LAN WAN CAN Telecommunications Firewall Load Balancing Traffic Shaping VoIP GRE IPSec IP Encryption MoIP NAT Data Capture ISO 9000 Information Assurance Security Engineering
10+ yrs exp

Senior Network Engineer

F5 Inc

Remote 60 days ago $179,500$269,300
BGP OSPF Python Ansible AWX Git CI/CD AWS Azure GCP JunOS EOS NX-OS IOS-XE TCP/IP VPN Firewalls DDoS Mitigation IDS/IPS NIST 800-53
10+ yrs exp Remote