AI Engineer V, LLM Gateway, FM Hosting

Capital One Financial

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
San Jose, CAMcLean, VACambridge, MANew York, NY
Salary
$229,900–$262,400 / yr
Posted
3 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $204k
This role $246k
$150k most similar roles pay here $274k

This role pays more than 76% of similar roles. Most pay $162,000–$246,150 — the shaded band above. At the midpoint, this role pays about $246k versus about $204k for comparable roles.

Based on 240 similar postings.

Employer

About Capital One Financial

Capital One Financial is a bank holding company specializing in credit cards, auto loans, banking, and savings products, known for its data-driven approach to consumer and commercial finance. Industry: Financial Services & Banking

Capital One Financial currently has 878 open roles on FindRole.

Listed pay typically runs $197,300–$225,100 across 872 roles with salary data.

Most-posted roles

View all roles at Capital One Financial

At a glance

TL;DR · AI Engineer V, LLM Gateway, FM Hosting

AI Engineer 5 (LLM Gateway, FM Hosting) joins the Intelligent Foundations and Experiences team to develop and deploy proprietary AI solutions that enhance products for millions of customers. This role involves designing, testing, and supporting software components including foundation model training, large language model inference, multi-agent workflows, similarity search, guardrails, and model evaluation. The engineer will optimize performance metrics like scalability, cost, latency, and throughput while leading cost-performance governance reviews and mentoring other engineers. Key technologies include Python, Go, Scala, CUDA, Java, C++, C#, PyTorch, Huggingface, VectorDBs, and AWS Ultraclusters. The work focuses on building high-performance AI infrastructure and multi-model orchestration pipelines to solve complex problems in the banking sector. The role requires a deep technical foundation in engineering and mathematics to implement state-of-the-art techniques for large-scale production systems while ensuring ethical deployment standards.

What does a AI Engineer earn in California?

Median $246150 from 73 postings across 11 companies.

See salary data

What you'll do

  • Develop and support AI software components including foundation model training, LLM inference, and multi-agent workflows.
  • Implement state-of-the-art optimization techniques to improve performance, scalability, cost, and latency of production AI systems.
  • Design and optimize multi-model orchestration pipelines integrating LLMs, vector search, and domain-specific models into unified systems.
  • Lead cost-performance governance reviews by tracking GPU utilization, model throughput, and inference cost efficiency.
  • Lead design councils to ensure technical consistency and compliance with established AI engineering standards.
  • Mentor senior engineers to foster cross-domain learning and improve organizational technical maturity.
  • Translate complex research papers into practical production techniques for large-scale AI infrastructure.

What we're looking for

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus 6 years of experience developing AI/ML algorithms.
  • Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus 4 years of experience developing AI/ML algorithms.
  • At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java.
  • Experience leading development of AI systems with trade-offs regarding cost, latency, throughput, and accuracy (preferred).
  • 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (preferred).
  • Experience designing, developing, delivering, and supporting complex AI systems (preferred).
  • Experience developing AI/ML algorithms using Python, C++, C#, Java, CUDA, or Golang (preferred).
  • Experience applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization and cost (preferred).

More like this

Similar roles

AI Engineer V

Capital One Financial

San Jose, CA +3 3 days ago $229,900$262,400
LLM Inference Python PyTorch Huggingface CUDA VectorDBs AWS Go Scala C++ C# Java Multi-agent Workflows Model Optimization Similarity Search
6+ yrs exp

AI Engineer V

Capital One Financial

San Jose, CA +3 3 days ago $229,900$262,400
LLM Python PyTorch CUDA Huggingface VectorDBs AWS Go Scala C++ Java C# Agentic AI Model Optimization Multi-agent Workflows Similarity Search Machine Learning
6+ yrs exp

AI Engineer V

Capital One Financial

Cambridge, MA +3 3 days ago $229,900$262,400
LLM Python PyTorch CUDA Huggingface VectorDBs AWS Go Scala C++ Java C# Agentic AI Model Optimization Multi-agent Workflows Similarity Search Machine Learning
6+ yrs exp

AI Engineer V

Capital One Financial

San Jose, CA +3 3 days ago $229,900$262,400
LLM Python PyTorch CUDA Huggingface VectorDBs AWS Go Scala C++ Java C# Agentic AI Model Optimization Multi-agent Workflows Similarity Search Machine Learning
6+ yrs exp

AI Engineer V, Gen AI Platform Services

Capital One Financial

San Francisco, CA +4 3 days ago $229,900$262,400
Python Go Scala CUDA C++ Java PyTorch Huggingface AWS VectorDBs LLM Inference Multi-agent Workflows Model Optimization Similarity Search SaaS AI Machine Learning
6+ yrs exp

AI Engineer V

Capital One Financial

San Jose, CA +3 3 days ago $229,900$262,400
Python Go CUDA PyTorch Huggingface AWS VectorDBs LLM Inference Multi-agent Workflows Similarity Search C++ Scala Java C# Model Optimization Machine Learning AI Engineering
6+ yrs exp