Browse tech roles

Basic role filtering by workplace, salary floor, and post age. For full AI matching and advanced filtering upload your resume using AI Match.

20 of up to 20 (filtered)

Product Manager, BioNeMo Inference

Nvidia

Santa Clara, CA 7 days ago $148,000$224,250
Actively hiring Confirmed live yesterday Posted this week High trust Competitive pay
NVIDIA NIM Triton TensorRT-LLM vLLM Ray KServe Kubeflow Kubernetes Docker Helm CI/CD MLOps GPU Infrastructure SDKs
5+ yrs exp

Senior Lead Software Engineer, Inference Platform Engineer

JPMorgan Chase

Seattle, WA 11 days ago
Actively hiring Confirmed live yesterday High trust
Kubernetes Docker Python Go Java C# Infrastructure as Code CI/CD Cloud Computing Microservices ML Ops MLflow Prometheus Grafana vLLM NVIDIA GPU Infrastructure Transformer Architecture SDLC
5+ yrs exp

Machine Learning Engineer, Causal Inference, Level V

Snap Inc.

Bellevue, WA +4 13 days ago $209,000$313,000
Actively hiring Confirmed live yesterday High trust Above market
Causal Inference Python A/B Testing Uplift Modeling scikit-learn NumPy pandas CausalML EconML DoWhy Machine Learning Statistics Bayesian Inference Experimentation Infrastructure
5+ yrs exp

Engineering Manager, LLM Inference & Deployment at Scale

Nvidia

Santa Clara, CA 14 days ago $224,000$356,500
Actively hiring Confirmed live yesterday High trust Above market
LLMs VLMs TensorRT TensorRT-LLM vLLM SGLang Quantization Speculative Decoding Continuous Batching Prefix Caching KV-cache Optimization Distributed Computing GPU Cluster Orchestration Model Serving Inference Optimization Deep Learning
8+ yrs exp Hybrid

AI Inference Engineer

F5 Inc

San Jose, CA +1 15 days ago $176,600$265,000
Actively hiring Confirmed live yesterday High trust Above market
vLLM TGI NVIDIA Triton Python C++ Rust TensorRT Llama.cpp Ollama Kubernetes Docker AWS GCP Azure CUDA TPUs MLOps
Hybrid

Senior Deep Learning Software Engineer, Inference

Nvidia

Remote (CA) +4 15 days ago $152,000$241,500
Actively hiring Confirmed live yesterday High trust Below market
CUDA C++ Python PyTorch vLLM SGLang Triton CUTLASS NCCL NVSHMEM Deep Learning LLM Generative AI GPU Programming Performance Optimization Profiling Agile
5+ yrs exp Remote

Senior Software Engineer, AI Inference Performance

Nvidia

Santa Clara, CA 16 days ago $184,000$287,500
Actively hiring Confirmed live yesterday High trust Above market
CUDA Python C++ Rust TensorRT-LLM vLLM SGLang Triton PyTorch NVIDIA Nsight Systems CUTLASS NCCL Quantization Distributed Systems GPU Architecture Model Parallelism Speculative Decoding
6+ yrs exp

Senior Applied Scientist Engineer, Training & Inference

Adobe

San Jose, CA 16 days ago $216,400$313,300
Actively hiring Confirmed live yesterday High trust Above market
PyTorch Python FSDP Tensor Parallelism Pipeline Parallelism vLLM TensorRT GPU Distributed Training Inference Optimization Diffusion Models Flow Matching Multi-node GPU Performance Profiling

Software Engineer, Inference (AI Data Engineering)

SpaceX

Palo Alto, CA 23 days ago $135,000$175,000
Actively hiring Confirmed live yesterday High trust Competitive pay
SGLang vLLM TensorRT-LLM Triton Rust C++ Python Go gRPC REST Docker Kubernetes PostgreSQL ClickHouse MongoDB CI/CD GPU Kernels Quantization Speculative Decoding
2+ yrs exp

Staff Machine Learning Scientist, Applied Causal Inference

DoorDash, Inc

San Francisco, CA +4 24 days ago $203,500$299,300
Actively hiring Confirmed live yesterday High trust Competitive pay
Causal Inference Econometrics Uplift Modeling Heterogeneous Treatment Effect Models Counterfactual Evaluation Doubly Robust Estimation Double ML Diff-in-Diff Synthetic Controls CUPED Contextual Bandits Off-policy Evaluation Machine Learning ML Engineering

Staff Machine Learning Engineer, Causal Inference

DoorDash, Inc

San Francisco, CA +4 24 days ago $203,500$299,300
Actively hiring Confirmed live yesterday High trust Competitive pay
Causal Inference Econometrics Uplift Modeling Double ML Off-policy Evaluation Contextual Bandits Doubly Robust Estimation Diff-in-Diff CUPED Machine Learning Experimentation

Staff Machine Learning Engineer, Diffusion, Generative Modeling and Inference

Snap Inc.

Bellevue, WA +4 27 days ago $229,000$343,000
Actively hiring Confirmed live yesterday High trust Above market
Generative AI Machine Learning Deep Learning Neural Networks Computer Vision PyTorch TensorFlow JAX MLX scikit-learn Diffusion Models Model Compression Quantization Distillation GPU CPU NPU
8+ yrs exp

Senior Data Scientist, Payments

Airbnb

Remote 28 days ago $179,000$210,000
Actively hiring Confirmed live yesterday High trust Competitive pay
Causal Inference Machine Learning Python R SQL LLM Statistical Modeling Experimentation Data Visualization Econometric Regression Data Science
5+ yrs exp Remote

Senior Machine Learning Engineer, Foundation Models Inference

Apple Inc

Seattle, WA 28 days ago $175,000$308,500
Actively hiring Confirmed live yesterday High trust Competitive pay
LLM Inference PyTorch JAX TensorFlow Python Go CUDA Triton Kubernetes Docker AWS GCP TensorRT-LLM vLLM SGLang TGI Nvidia Triton Server Transformers
5+ yrs exp

Senior Machine Learning Engineer, Foundation Models Inference

Apple Inc

Santa Clara, CA 28 days ago $184,700$324,800
Actively hiring Confirmed live yesterday High trust Above market
LLM Inference PyTorch JAX TensorFlow Python Go CUDA Triton Kubernetes Docker AWS GCP TensorRT-LLM vLLM SGLang TGI Nvidia Triton Server Transformers
5+ yrs exp

Senior Product Manager, AI Inference Performance

Nvidia

Santa Clara, CA 29 days ago $208,000$327,750
Actively hiring Confirmed live yesterday High trust Above market
AI Inference TensorRT-LLM vLLM SGLang NVIDIA Dynamo Triton Inference Server Quantization Speculative Decoding KV Caching LLM Deep Learning Generative AI Benchmarking Capacity Planning Autoscaling
10+ yrs exp

Senior AI Inference Platform Engineer

Apple Inc

Seattle, WA 30 days ago $175,000$308,500
Actively hiring Confirmed live 2 days ago High trust Competitive pay
Python Go C++ Triton TensorRT-LLM vLLM Kubernetes Prometheus Grafana Splunk CI/CD Data Pipelines Performance Benchmarking Distributed Systems Capacity Planning
7+ yrs exp

Senior AI Inference Platform Engineer

Apple Inc

Seattle, WA 31 days ago $175,000$308,500
Actively hiring Confirmed live yesterday High trust Competitive pay
Python Go C++ Triton TensorRT-LLM vLLM Kubernetes Prometheus Grafana Splunk CI/CD Data Pipelines Performance Benchmarking Distributed Systems Capacity Planning
7+ yrs exp

Staff Applied Scientist

Adobe

San Jose, CA 32 days ago $216,400$313,300
Actively hiring Confirmed live yesterday High trust Above market
LLM VLM Python PyTorch vLLM TensorRT-LLM SGLang Triton Inference Server TGI Kubernetes Ray Spark Dask Distributed Systems Quantization Vector Search

Senior Data Scientist, Trust (Inference)

Airbnb

Remote 35 days ago $179,000$210,000
Actively hiring Confirmed live yesterday High trust Competitive pay
Causal Inference Bayesian Modeling Experimentation Machine Learning Python R SQL Data Visualization A/B Testing Quasi-experiments Data Science
5+ yrs exp Remote