Browse tech roles

Basic role filtering by workplace, salary floor, and post age. For full AI matching and advanced filtering upload your resume using AI Match.

16 of up to 20 (filtered)

AI Inference Engineer

F5 Inc

San Jose, CA +1 15 days ago $176,600$265,000
Actively hiring Confirmed live yesterday High trust Above market
vLLM TGI NVIDIA Triton Python C++ Rust TensorRT Llama.cpp Ollama Kubernetes Docker AWS GCP Azure CUDA TPUs MLOps
Hybrid

Senior Software Engineer, AI Inference Performance

Nvidia

Santa Clara, CA 16 days ago $184,000$287,500
Actively hiring Confirmed live 2 days ago High trust Above market
CUDA Python C++ Rust TensorRT-LLM vLLM SGLang Triton PyTorch NVIDIA Nsight Systems CUTLASS NCCL Quantization Distributed Systems GPU Architecture Model Parallelism Speculative Decoding
6+ yrs exp

Senior Applied Scientist Engineer, Training & Inference

Adobe

San Jose, CA 16 days ago $216,400$313,300
Actively hiring Confirmed live 2 days ago High trust Above market
PyTorch Python FSDP Tensor Parallelism Pipeline Parallelism vLLM TensorRT GPU Distributed Training Inference Optimization Diffusion Models Flow Matching Multi-node GPU Performance Profiling

Software Engineer, Inference (AI Data Engineering)

SpaceX

Palo Alto, CA 23 days ago $135,000$175,000
Actively hiring Confirmed live 2 days ago High trust Competitive pay
SGLang vLLM TensorRT-LLM Triton Rust C++ Python Go gRPC REST Docker Kubernetes PostgreSQL ClickHouse MongoDB CI/CD GPU Kernels Quantization Speculative Decoding
2+ yrs exp

Senior AI Inference Platform Engineer

Apple Inc

Seattle, WA 30 days ago $175,000$308,500
Actively hiring Confirmed live 2 days ago High trust Competitive pay
Python Go C++ Triton TensorRT-LLM vLLM Kubernetes Prometheus Grafana Splunk CI/CD Data Pipelines Performance Benchmarking Distributed Systems Capacity Planning
7+ yrs exp

Senior AI Inference Platform Engineer

Apple Inc

Seattle, WA 31 days ago $175,000$308,500
Actively hiring Confirmed live yesterday High trust Competitive pay
Python Go C++ Triton TensorRT-LLM vLLM Kubernetes Prometheus Grafana Splunk CI/CD Data Pipelines Performance Benchmarking Distributed Systems Capacity Planning
7+ yrs exp

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

McLean, VA +2 39 days ago $229,900$262,400
Actively hiring Confirmed live 2 days ago Trusted Above market
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Scala Java C# Machine Learning Similarity Search Model Evaluation Observability Optimization Techniques
6+ yrs exp

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 51 days ago $229,900$262,400
Actively hiring Confirmed live 2 days ago Low trust Above market
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning Similarity Search Model Evaluation Observability
6+ yrs exp

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 71 days ago $229,900$262,400
Actively hiring Confirmed live yesterday Low trust Above market
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning Similarity Search Model Evaluation Observability Optimization Techniques
6+ yrs exp

Senior Lead AI Engineer

Capital One Financial

San Jose, CA +4 100 days ago $229,900$262,400
Actively hiring Confirmed live 2 days ago Low trust Above market
Python Go Scala Java C++ C# PyTorch Huggingface AWS VectorDBs LLM Inference Nemo Guardrails Machine Learning Similarity Search Inference Optimization SaaS AI
6+ yrs exp

Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 102 days ago $197,300$225,100
Actively hiring Confirmed live 2 days ago Low trust Competitive pay
LLM Inference Python PyTorch Huggingface VectorDBs AWS Nemo Guardrails Go Scala Java C++ C# Similarity Search Machine Learning Model Evaluation Observability
4+ yrs exp

Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 102 days ago $197,300$225,100
Actively hiring Confirmed live 2 days ago Low trust Competitive pay
LLM Inference Python PyTorch Huggingface VectorDBs AWS Nemo Guardrails Go Scala Java C++ C# Similarity Search Machine Learning Model Evaluation Observability
4+ yrs exp

Engineering Manager, Inference Benchmarking

Nvidia

Remote (Santa Clara, CA) +1 106 days ago $224,000$356,500
Actively hiring Confirmed live yesterday High trust Above market
LLM vLLM TRT-LLM SGLang Kubernetes Prometheus ZMQ DCGM PyNVML Helm Distributed Systems Microservices open-source Inference Infrastructure Benchmarking Computer Vision
8+ yrs exp Remote

Senior Software Engineer, AI Inference Systems

Nvidia

Santa Clara, CA 136 days ago $184,000$287,500
Actively hiring Confirmed live yesterday High trust Above market
vLLM SGLang CUDA Python C++ PyTorch Triton MLIR LLVM XLA Docker Kubernetes Slurm NCCL Nsight Systems CI/CD AWS GCP Azure High-Performance Computing
7+ yrs exp Hybrid

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 162 days ago $229,900$262,400
Actively hiring Confirmed live yesterday Low trust Above market
LLM Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning LLM Inference Similarity Search Model Evaluation Observability
6+ yrs exp

Principal GenAI Inference Optimization Engineer

Amd

San Jose, CA 167 days ago $240,000$360,000
Actively hiring Confirmed live yesterday High trust Above market
GenAI LLM GPU Architecture vLLM SGLang Triton TensorRT-LLM PyTorch JAX TensorFlow Python C++ CUDA HIP Quantization Distributed Systems RDMA Profiling Benchmarking
Hybrid