AI Engineer 4

Capital One Financial

Confirmed live today High trust

Quick summary

Work type
On-site
Location
San Jose, CASan Francisco, CAMcLean, VACambridge, MANew York, NY
Salary
$197,300–$225,100 / yr
Employment
Full-time
Posted
4 days ago
Freshness
Confirmed live today

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $203k
This role $211k
$187k $230k
below market most similar roles pay here above market

This role pays more than 75% of similar roles. Most pay $192,050–$214,000 — the blue band above. At the midpoint, this role pays about $211k versus about $203k for comparable roles.

Based on 240 similar postings.

Employer

About Capital One Financial

Capital One Financial is a bank holding company specializing in credit cards, auto loans, banking, and savings products, known for its data-driven approach to consumer and commercial finance. Industry: Financial Services & Banking

Capital One Financial currently has 1917 open roles on FindRole.

Listed pay typically runs $197,300–$225,100 across 1045 roles with salary data.

Most-posted roles

View all roles at Capital One Financial

At a glance

TL;DR · AI Engineer 4

The AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) joins the Intelligent Foundations and Experiences team to develop proprietary AI solutions for banking. You will design, develop, test, deploy, and support AI software components, including foundation model training, large language model inference, agents, multi-agent workflows, similarity search, guardrails, and model evaluation. The role involves inventing optimization techniques to improve performance, scalability, cost, and latency for production systems while defining service-level objectives for reliability. You will utilize a technical stack including AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, Python, Go, Scala, CUDA, and Java. This position solves complex problems regarding model governance, observability, and ethical alignment, ensuring that scalable, high-performance AI infrastructure delivers reliable and responsible customer experiences across the enterprise.

What does a AI Engineer earn in California?

Median $246150 from 84 postings across 12 companies.

See salary data

What you'll do

  • Design, develop, test, deploy, and support AI software components including foundation model training and LLM inference.
  • Build and manage agents, multi-agent workflows, similarity search, guardrails, and model evaluation systems.
  • Implement foundation model optimization techniques to improve scalability, cost, latency, and throughput of production systems.
  • Own the end-to-end architecture for complex AI systems to ensure maintainability, observability, and ethical alignment.
  • Define and maintain service-level objectives for AI reliability, including latency, uptime, and model performance drift.
  • Optimize GPU/TPU utilization and accelerate model inference pipelines in collaboration with infrastructure engineering.
  • Lead cross-functional technical reviews to ensure new AI deployments meet security, data governance, and compliance standards.
  • Mentor senior associates on scalable design, performance tuning, and translating research into production environments.

What we're looking for

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus 4 years of experience developing AI and ML algorithms.
  • Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus 2 years of experience developing AI and ML algorithms.
  • At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java.
  • Experience leading development of AI systems with tradeoff decisions around cost, latency, throughput, and accuracy (preferred).
  • 6+ years of experience deploying scalable and responsible AI solutions on cloud platforms (preferred).
  • Experience designing, developing, delivering, and supporting AI services (preferred).
  • Experience developing AI and ML algorithms using Python, C++, C#, Java, CUDA, or Golang (preferred).
  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software (preferred).

More like this

Similar roles

AI Engineer 4, Gen AI Platform Services

Capital One Financial

San Francisco, CA +4 5 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS VectorDBs LLM Inference Agentic AI Distributed Systems Similarity Search GPU TPU Model Evaluation Observability
4+ yrs exp

AI Engineer 4

Capital One Financial

New York, NY +4 13 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS Google Cloud Azure VectorDBs LLM Inference Distributed Systems Model Evaluation Observability GPU TPU
4+ yrs exp

AI Engineer 4

Capital One Financial

McLean, VA +4 19 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS Google Cloud Azure VectorDBs LLM Inference Agentic AI Distributed Systems Model Evaluation Observability
4+ yrs exp

AI Engineer 4

Capital One Financial

San Jose, CA +4 19 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS Google Cloud Azure VectorDBs LLM Inference Agentic AI Distributed Systems Model Evaluation Observability
4+ yrs exp

AI Engineer 4

Capital One Financial

San Jose, CA +4 23 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS Google Cloud Azure VectorDBs LLM Inference Agentic AI Distributed Systems Model Evaluation Observability
4+ yrs exp

AI Engineer 4

Capital One Financial

San Jose, CA +4 23 days ago $197,300–$225,100
Python Go Scala CUDA Java C++ C# PyTorch Huggingface AWS Google Cloud Azure VectorDBs LLM Inference Agentic AI Distributed Systems Model Evaluation Observability
4+ yrs exp