Browse tech roles

Basic role filtering by workplace, salary floor, and post age. For full AI matching and advanced filtering upload your resume using AI Match.

14 of up to 20 (filtered)

Senior Lead Software Engineer, Inference Platform Engineer

JPMorgan Chase

Seattle, WA 11 days ago
Actively hiring Confirmed live yesterday High trust
Kubernetes Docker Python Go Java C# Infrastructure as Code CI/CD Cloud Computing Microservices ML Ops MLflow Prometheus Grafana vLLM NVIDIA GPU Infrastructure Transformer Architecture SDLC
5+ yrs exp

Senior Deep Learning Software Engineer, Inference

Nvidia

Remote (CA) +4 15 days ago $152,000$241,500
Actively hiring Confirmed live 2 days ago High trust Below market
CUDA C++ Python PyTorch vLLM SGLang Triton CUTLASS NCCL NVSHMEM Deep Learning LLM Generative AI GPU Programming Performance Optimization Profiling Agile
5+ yrs exp Remote

Senior Software Engineer, AI Inference Performance

Nvidia

Santa Clara, CA 16 days ago $184,000$287,500
Actively hiring Confirmed live yesterday High trust Above market
CUDA Python C++ Rust TensorRT-LLM vLLM SGLang Triton PyTorch NVIDIA Nsight Systems CUTLASS NCCL Quantization Distributed Systems GPU Architecture Model Parallelism Speculative Decoding
6+ yrs exp

Senior Applied Scientist Engineer, Training & Inference

Adobe

San Jose, CA 16 days ago $216,400$313,300
Actively hiring Confirmed live 2 days ago High trust Above market
PyTorch Python FSDP Tensor Parallelism Pipeline Parallelism vLLM TensorRT GPU Distributed Training Inference Optimization Diffusion Models Flow Matching Multi-node GPU Performance Profiling

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

McLean, VA +2 39 days ago $229,900$262,400
Actively hiring Confirmed live 2 days ago Trusted Above market
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Scala Java C# Machine Learning Similarity Search Model Evaluation Observability Optimization Techniques
6+ yrs exp

Senior System Software Engineer, Agentic Inference

Nvidia

Santa Clara, CA 46 days ago $224,000$356,500
Actively hiring Confirmed live 2 days ago High trust Above market
Rust Python vLLM SGLang TensorRT-LLM GPU Deep Learning Generative AI Performance Analysis NIXL Open Source Software high-performance networking
10+ yrs exp Hybrid

Senior Inference Engineer, GPU Kernel Optimization

Nvidia

Santa Clara, CA +3 46 days ago $184,000$287,500
Actively hiring Confirmed live yesterday High trust Competitive pay
CUDA Triton CUTLASS Python C++ TRT-LLM SGLang vLLM CUPTI NSYS NCU PTX SASS LLVM MLIR FlashInfer Kernel Optimization Agentic AI Systems
6+ yrs exp

Senior GPU Inference Performance Engineer

Amd

Santa Clara, CA 64 days ago $164,000$246,000
Actively hiring Confirmed live 2 days ago High trust Competitive pay
ROCm CUDA Python C++ Kubernetes Docker RDMA RoCE NCCL RCCL HIP Linux HPC MPI UCX
Hybrid

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 71 days ago $229,900$262,400
Actively hiring Confirmed live yesterday Low trust Above market
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning Similarity Search Model Evaluation Observability Optimization Techniques
6+ yrs exp

Senior System Software Engineer, Dynamo-Triton Inference Server

Nvidia

Remote (Santa Clara, CA) +1 71 days ago $152,000$241,500
Actively hiring Confirmed live yesterday High trust Competitive pay
Rust C++ Python Triton Inference Server TensorRT PyTorch ONNX vLLM TRT-LLM GPU Distributed Systems High-Performance Networking Open Source GitHub Agile
5+ yrs exp Remote

Senior DL Algorithms Engineer, Inference Performance

Nvidia

Remote (Canada) 129 days ago $152,000$241,500
Actively hiring Confirmed live 2 days ago High trust Below market
Deep Learning LLM GPU Architecture PyTorch TRT-LLM vLLM SGLang FlashInfer CUDA OpenCL Performance Profiling HPC Neural Networks
3+ yrs exp Remote

Senior Staff Engineer, ML Inference

Shopify

Remote 133 days ago
Confirmed live 2 days ago High trust
ML Inference CUDA TensorRT Triton TVM Python C++ DeepSpeed ONNX GPU Pruning Distillation Batching Distributed Computing MLOps Monitoring Observability
Remote

Senior Software Engineer, AI Inference Systems

Nvidia

Santa Clara, CA 136 days ago $184,000$287,500
Actively hiring Confirmed live yesterday High trust Above market
vLLM SGLang CUDA Python C++ PyTorch Triton MLIR LLVM XLA Docker Kubernetes Slurm NCCL Nsight Systems CI/CD AWS GCP Azure High-Performance Computing
7+ yrs exp Hybrid

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 162 days ago $229,900$262,400
Actively hiring Confirmed live yesterday Low trust Above market
LLM Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning LLM Inference Similarity Search Model Evaluation Observability
6+ yrs exp