Senior Software Engineer, AI Inference Systems
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 72% of similar roles. Most pay $158,400–$235,750 — the shaded band above. At the midpoint, this role pays about $221k versus about $197k for comparable roles.
Based on 240 similar postings.
Employer
F5, Inc. is an American technology company specializing in application security, multi-cloud management, online fraud prevention, application delivery networking, application availability and performance, and network security, access, and authorization.
F5 Inc currently has 25 open roles on FindRole.
Listed pay typically runs $152,200–$228,400 across 25 roles with salary data.
Most-posted roles
At a glance
The AI Inference Engineer joins the team to bridge the gap between high-performance model development and optimized deployment environments. This role focuses on optimizing Large Language Models for inference across diverse settings, from GPU-rich data centers to resource-constrained edge devices. The engineer will build and maintain robust inference engines using vLLM, TGI, and NVIDIA Triton while designing auto-scaling architectures via Kubernetes for real-time and batch pipelines. Key responsibilities include hardware acceleration for NVIDIA GPUs, Apple Silicon, TPUs, and LPUs, alongside establishing observability frameworks to monitor metrics like Time to First Token and memory bandwidth. Required skills include proficiency in Python, C++, Rust, or Golang, along with experience in Docker, cloud platforms like AWS, GCP, and Azure. The role solves the technical challenge of ensuring enterprise-grade reliability and low-latency performance for large-scale AI applications.
Skills
What you'll do
What we're looking for
More like this
Nvidia
Apple Inc
Apple Inc
Amd
SpaceX
Apple Inc