Manager, Applied AI Science, LLM/VLM-as-judge

Apple Inc

Confirmed live today High trust

Quick summary

Work type
On-site
Location
Cupertino, CA
Salary
$206,200–$356,400 / yr
Posted
5 days ago
Freshness
Confirmed live today

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $223k
This role $281k
$153k $378k
below market most similar roles pay here above market

This role pays more than 88% of similar roles. Most pay $192,050–$254,750 — the blue band above. At the midpoint, this role pays about $281k versus about $223k for comparable roles.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 3595 open roles on FindRole.

Listed pay typically runs $166,600–$277,600 across 2772 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Manager, Applied AI Science, LLM/VLM-as-judge

AIML - Manager, Applied AI Science - LLM/VLM-as-judge will lead a team of applied AI scientists and machine learning engineers within the Evaluation organization. The manager is responsible for setting the technical vision and roadmap for autograder development, hiring and developing talent, and delivering reliable, scalable LLM/VLM-based evaluation systems. Day-to-day tasks include designing state-of-the-art autograders, standardizing end-to-end evaluation workflows, and navigating complex cross-functional dependencies. The role requires deep expertise in GenAI models, model alignment, benchmark creation, rubric design, and human-model alignment. The work focuses on providing principled assessments for diverse features like Search and Siri by building trustworthy automated evaluation systems that serve as a foundation for product-quality decisions across the company's internal generative AI products.

What you'll do

  • Lead and grow a high-performing team of applied AI scientists and machine learning engineers.
  • Set the technical vision and roadmap for autograder development and evaluation methodologies.
  • Deliver reliable and scalable LLM/VLM-based evaluation systems for product-quality decisions.
  • Standardize and automate end-to-end evaluation workflows to enable organizational scaling.
  • Translate emerging GenAI research into reliable, production-ready systems and frameworks.
  • Manage cross-functional dependencies and align diverse stakeholders on AI evaluation strategy.
  • Design rubrics, evaluation sets, and validation processes for human-model alignment.

What we're looking for

  • MS/PhD degree in Computer Science, Machine Learning, AI, or a related field.
  • 1+ years of industry experience building LLM/VLM-based evaluators, autograders, or automated evaluation systems.
  • 5+ years of hands-on experience in machine learning, model alignment, or benchmark creation.
  • Track record of managing teams of 3+ machine learning engineers and/or AI scientists.
  • Deep understanding of GenAI models, evaluation methodologies, rubric design, and human-model alignment.
  • Strong technical judgment and a track record of leading complex, ambiguous AI/ML projects from definition through delivery (preferred).
  • Experience translating emerging GenAI techniques and research into reliable, production-ready systems (preferred).
  • Experience scaling AI/ML solutions into reusable platforms, frameworks, or standardized workflows (preferred).

More like this

Similar roles

Manager, Applied AI Science, LLM/VLM-as-judge

Apple Inc

Cupertino, CA 22 days ago $206,200–$356,400
LLM VLM GenAI Machine Learning Model Alignment Benchmark Creation Evaluation Methodologies Rubric Design Validation Calibration Human-Model Alignment Production-ready Systems
5+ yrs exp

Machine Learning Engineer, AI Evaluation & LLM Systems

Apple Inc

Cupertino, CA 73 days ago $150,400–$225,300
Python C++ PyTorch TensorFlow JAX Large Language Models Multimodal AI Generative AI CI/CD Git Distributed Computing Data Processing Model Evaluation Statistical Analysis Supervised Learning
1+ yrs exp