Manager, AI Engineering

Mastercard

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
O Fallon, MO
Salary
$140,000–$231,000 / yr
Posted
8 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $212k
This role $186k
$126k most similar roles pay here $267k

This role pays less than 65% of similar roles. Most pay $170,300–$253,700 — the shaded band above. At the midpoint, this role pays about $186k versus about $212k for comparable roles.

Based on 240 similar postings.

Employer

About Mastercard

Mastercard is a global technology company in the payments industry, processing transactions between financial institutions and merchants using its extensive network of credit, debit, and prepaid card products. Industry: Payments Technology & Financial Services

Mastercard currently has 116 open roles on FindRole.

Listed pay typically runs $122,000–$207,000 across 104 roles with salary data.

Most-posted roles

View all roles at Mastercard

At a glance

TL;DR · Manager, AI Engineering

Manager, AI Engineering (Tester ) joins the Business & Market Insights group to lead quality engineering for the Operational Intelligence Program. This role focuses on ensuring Generative AI, LLM, and agentic systems are accurate, safe, and enterprise-ready. You will design end-to-end evaluation frameworks, build test suites for agentic workflows, develop RAG pipeline evaluations, and conduct structured red-teaming against prompt injections and data leakage. The position involves performing fairness audits, monitoring inference performance, and establishing quality gates within CI/CD pipelines. Required skills include Python, SQL, and experience with tools like RAGAS, TruLens, DeepEval, LangSmith, and PromptFlow. You will work with the OpenAI, Anthropic, Hugging Face, and LangChain ecosystems while utilizing Databricks, AWS, Grafana, and Datadog to monitor production systems. The role addresses critical challenges in hallucination detection, retrieval precision, and automated quality standards for complex AI systems.

What you'll do

  • Design and own end-to-end LLM evaluation frameworks including automated prompt regression, scoring, and hallucination detection.
  • Build comprehensive test suites for agentic AI systems to validate tool selection, coordination, and multi-step reasoning workflows.
  • Develop RAG pipeline evaluation frameworks to assess retrieval precision, context faithfulness, and answer grounding.
  • Lead structured red-teaming and adversarial testing to defend against prompt injection, jailbreaks, and data leakage.
  • Execute fairness, bias, and Responsible AI audits to ensure demographic equity and explainability.
  • Perform inference performance benchmarking and enforce LLM quality gates within CI/CD pipelines on Databricks.
  • Build production monitoring and drift detection pipelines using observability tools like Grafana and Datadog.
  • Define the AI testing roadmap, quality standards, and evaluation metrics for all Generative AI workstreams.

What we're looking for

  • Bachelor's or Master's degree in Computer Science, AI/ML, or Software Engineering.
  • Experience leading AI/ML quality engineering or LLM testing programs in production environments.
  • Expertise in testing LLM and Gen AI systems including prompt testing, hallucination detection, and RAG pipeline assessment.
  • Proficiency with evaluation tools such as RAGAS, DeepEval, TruLens, LangSmith, PromptFlow, or Weights & Biases Evals.
  • Strong Python programming skills for building test automation scripts and SQL proficiency.
  • Knowledge of LLM ecosystems including OpenAI, Anthropic, Hugging Face, and LangChain/LangGraph.
  • Experience with MLOps/LLMOps pipelines (MLflow, Databricks, SageMaker) and integrating quality gates into CI/CD workflows.
  • Familiarity with cloud AI infrastructure (AWS, Azure, or GCP) and observability tooling like Grafana, Datadog, or CloudWatch.

More like this

Similar roles

Manager, AI Engineering

Rockwell Automation

Milwaukee, WI 51 days ago
Python Azure AI GenAI LLM RAG Vector Databases MLOps LLMOps CI/CD Microsoft 365 SharePoint Graph API Power Platform GitHub Actions API Data Engineering
10+ yrs exp Hybrid

Manager, AI Evaluation Engineering

Caterpillar

Broomfield, CO +3 10 days ago $147,760$240,110
Generative AI LLMs SLMs RAG Prompt Engineering LangChain LangGraph Vector Databases Azure AWS GCP SageMaker Bedrock Snowflake Cortex CI/CD Automated Testing LoRA A/B Testing Feature Flags

Manager, AI Evaluation Engineering

Caterpillar

Chicago, IL +3 10 days ago $147,760$240,110
Generative AI LLMs SLMs RAG LangChain LangGraph Prompt Engineering Vector Databases Azure AWS GCP SageMaker Bedrock Snowflake Cortex CI/CD Automated Testing LoRA Feature Flags

Senior AI Engineer

Mastercard

O Fallon, MO 11 days ago $115,000$184,000
Generative AI LLM RAG LangGraph CrewAI AutoGen Python PyTorch TensorFlow Databricks MLflow Pinecone Neo4j AWS FastAPI SQL Hugging Face Prompt Engineering LoRA PEFT

Manager, AI Engineering

Caterpillar

Chicago, IL +2 10 days ago $147,760$240,110
Generative AI LLMs RAG LangChain LangGraph Semantic Kernel CrewAI Python Go Java AWS Azure Google Cloud CI/CD infrastructure-as-code Agile Azure DevOps Prompt Engineering

Manager, AI Engineering

Caterpillar

Chicago, IL +2 10 days ago $147,760$240,110
Generative AI LLMs RAG LangChain LangGraph Semantic Kernel CrewAI Python Go Java AWS Azure Google Cloud CI/CD Infrastructure-as-Code Agile Azure DevOps Prompt Engineering