AI Systems Performance Engineer

Broadcom

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
San Jose, CA
Salary
$141,300–$226,000 / yr
Posted
144 days ago
Freshness
Confirmed live yesterday
Closes
Oct 17, 2026

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $197k
This role $184k
$130k most similar roles pay here $242k

This role pays less than 55% of similar roles. Most pay $162,000–$231,362 — the shaded band above. At the midpoint, this role pays about $184k versus about $197k for comparable roles.

Based on 240 similar postings.

Employer

About Broadcom

Broadcom is a global semiconductor and infrastructure software company that designs and markets a wide range of networking, storage, and wireless connectivity solutions. Industry: Semiconductors & Infrastructure Software

Broadcom currently has 119 open roles on FindRole.

Listed pay typically runs $121,900–$195,000 across 105 roles with salary data.

Most-posted roles

View all roles at Broadcom

At a glance

TL;DR · AI Systems Performance Engineer

As an AI Systems Performance Engineer within the Performance Lab, you will drive performance benchmarking for AI inference, training, and storage workloads. You will focus on optimizing the Ethernet fabric connecting servers to ensure seamless data flow for distributed AI workloads. Your daily responsibilities include executing rigorous benchmarks like MLPerf and NCCL tests, isolating complex system bottlenecks across Linux OS, server hardware, and Ethernet switches, and developing automated testing frameworks. To succeed, you must possess deep expertise in Linux systems, Python, C++, and PyTorch, while demonstrating a strong understanding of how AI models consume compute and network resources. You will also utilize RDMA, RoCEv2, Docker, and Kubernetes to troubleshoot distributed systems. This role solves critical performance puzzles by tuning network parameters to achieve maximum throughput and minimum latency for large-scale machine learning infrastructure.

What you'll do

  • Execute industry-standard AI performance benchmarks including MLPerf (Training and Inference) and NCCL tests.
  • Optimize Ethernet fabric parameters to ensure maximum throughput and minimum latency for distributed AI workloads.
  • Identify, isolate, and troubleshoot complex system bottlenecks across Linux OS, server hardware, and Ethernet switches.
  • Develop and implement robust automated testing frameworks and tools to streamline continuous benchmarking.
  • Generate performance reports to assist customers with deployment and marketing teams with product positioning.
  • Document test methodologies and provide actionable improvement recommendations to hardware, software, and networking stakeholders.

What we're looking for

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • 10 to 12+ years of related industry experience.
  • Deep familiarity and hands-on experience with Linux operating systems and system-level performance tuning.
  • Strong proficiency in programming and scripting languages, specifically Python and C++.
  • Familiarity with machine learning frameworks like PyTorch and understanding of AI model resource consumption.
  • Proven experience in performance testing and validating Ethernet switch systems.
  • Experience with RDMA (Remote Direct Memory Access) and RoCEv2.
  • Familiarity with containerization and orchestration tools such as Docker and Kubernetes.

More like this

Similar roles

Senior System Software Engineer, GPU Performance

Nvidia

Remote 6 days ago $152,000$241,500
HPC C++ Python CUDA NCCL UCX NVSHMEM MPI Infiniband Ethernet RDMA PyTorch TensorFlow Kubernetes Docker SLURM Ansible Performance Engineering Parallel Programming
Remote

Principal Developer, AI Networking

Nvidia

Remote (Santa Clara, CA) +3 91 days ago $272,000$431,250
NCCL RDMA CUDA PyTorch TensorFlow C++ Python Bash RoCE MPI SHARP Distributed Systems LLM Training High-Performance Networking Profiling Benchmarking
10+ yrs exp Remote

Senior AI/ML Platform Engineer

Amd

Santa Clara, CA 51 days ago $204,000$306,000
Python C++ Go Rust Kubernetes Ray Slurm vLLM SGLang Triton PyTorch JAX ROCm HIP CUDA MLflow Weights & Biases CI/CD Distributed Systems GPU Infrastructure
Hybrid

System Performance Engineer

Apple Inc

Cupertino, CA 2 days ago $129,300$225,300
C C++ Python Performance Analysis System Architecture Firmware OS Drivers Thermal Management Silicon CPU GPU DRAM Storage
2+ yrs exp