Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

Confirmed live 2 days ago Trusted

Quick summary

Work type
On-site
Location
McLean, VANew York, NYSan Jose, CA
Salary
$229,900–$262,400 / yr
Posted
39 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $219k
This role $246k
$175k most similar roles pay here $272k

This role pays more than 80% of similar roles. Most pay $191,000–$246,150 — the shaded band above. At the midpoint, this role pays about $246k versus about $219k for comparable roles.

Based on 240 similar postings.

Employer

About Capital One Financial

Capital One Financial is a bank holding company specializing in credit cards, auto loans, banking, and savings products, known for its data-driven approach to consumer and commercial finance. Industry: Financial Services & Banking

Capital One Financial currently has 998 open roles on FindRole.

Listed pay typically runs $197,300–$225,100 across 992 roles with salary data.

Most-posted roles

View all roles at Capital One Financial

At a glance

TL;DR · Senior Lead AI Engineer, FM Hosting, LLM Inference

The Senior Lead AI Engineer (FM Hosting, LLM Inference) joins the Intelligent Foundations and Experiences team to develop and deploy proprietary solutions that power core business functions. You will collaborate with cross-functional teams of engineers, scientists, and product managers to design, test, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, and observability. The role involves implementing state-of-the-art LLM optimization techniques to improve performance metrics like scalability, cost, latency, and throughput for production systems. You will utilize a technical stack featuring PyTorch, Huggingface, VectorDBs, Nemo Guardrails, and AWS Ultraclusters. Required skills include proficiency in Python, Go, Scala, or Java, along with expertise in hardware and software optimization to solve complex problems within the domain of large-scale production AI infrastructure for banking services.

What you'll do

  • Design, develop, test, and deploy AI software components including foundation model training and LLM inference.
  • Implement similarity search, guardrails, model evaluation, and observability for production AI systems.
  • Develop state-of-the-art LLM optimization techniques to improve performance, scalability, cost, latency, and throughput.
  • Utilize a broad stack of technologies including AWS Ultraclusters, Huggingface, VectorDBs, and PyTorch.
  • Contribute to the technical vision and long-term roadmap of foundational AI systems.
  • Translate complex scientific research into practical, high-performance production features.
  • Lead and mentor an engineering team while influencing cross-functional stakeholders.

What we're looking for

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields with at least 6 years of experience developing AI/ML algorithms.
  • Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields with at least 4 years of experience developing AI/ML algorithms.
  • At least 6 years of experience programming with Python, Go, Scala, or Java.
  • Experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, or Azure.
  • Experience designing, developing, integrating, delivering, and supporting complex AI systems including LLM inference and vector databases.
  • Experience applying state-of-the-art techniques to optimize training and inference software for hardware utilization, latency, throughput, and cost.
  • Demonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholders.
  • Excellent communication skills to articulate complex AI concepts to peers and present findings clearly.

More like this

Similar roles

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 71 days ago $229,900$262,400
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning Similarity Search Model Evaluation Observability Optimization Techniques
6+ yrs exp

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 162 days ago $229,900$262,400
LLM Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning LLM Inference Similarity Search Model Evaluation Observability
6+ yrs exp

Senior Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 51 days ago $229,900$262,400
LLM Inference Python Go PyTorch Huggingface AWS VectorDBs Nemo Guardrails C++ Java Scala C# Machine Learning Similarity Search Model Evaluation Observability
6+ yrs exp

Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 102 days ago $197,300$225,100
LLM Inference Python PyTorch Huggingface VectorDBs AWS Nemo Guardrails Go Scala Java C++ C# Similarity Search Machine Learning Model Evaluation Observability
4+ yrs exp

Lead AI Engineer, FM Hosting, LLM Inference

Capital One Financial

New York, NY +3 102 days ago $197,300$225,100
LLM Inference Python PyTorch Huggingface VectorDBs AWS Nemo Guardrails Go Scala Java C++ C# Similarity Search Machine Learning Model Evaluation Observability
4+ yrs exp

Senior Lead AI Engineer, LLM Gateway, FM Hosting

Capital One Financial

San Jose, CA +3 8 days ago $229,900$262,400
LLM Python Go Scala Java C++ C# AWS Huggingface PyTorch VectorDBs Nemo Guardrails Machine Learning LLM Inference Similarity Search Model Evaluation Observability
6+ yrs exp