Machine Learning Engineering Manager, Evaluation, Agentic Search Capabilities, Proactive

Apple Inc

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Santa Clara, CA
Salary
$206,200–$356,400 / yr
Posted
73 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $222k
This role $281k
$153k most similar roles pay here $378k

This role pays more than 85% of similar roles. Most pay $189,750–$254,750 — the shaded band above. At the midpoint, this role pays about $281k versus about $222k for comparable roles.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 2925 open roles on FindRole.

Listed pay typically runs $166,600–$277,600 across 2305 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Machine Learning Engineering Manager, Evaluation, Agentic Search Capabilities, Proactive

As the Machine Learning Engineering Manager, Evaluation, Agentic Search Capabilities, Proactive, you will join the Siri team to advance Apple Intelligence through personalized search and result ranking. You will lead a team of applied machine learning and software engineers to design and build scalable, automated systems for end-to-end evaluation of personalized Agentic Search. Your daily work involves developing data generation methods, creating metrics frameworks for quality assessment, and defining the long-term technical vision for Personal Question Answering. This role focuses on solving the challenge of providing low-latency production services that allow users to query personal information like emails and messages while maintaining privacy. You will utilize Python or C/C++ and apply knowledge of neural network architectures and machine learning fundamentals to improve search quality, data quality, and offline and online metrics for the product.

What you'll do

  • Lead a team of machine learning and software engineers to build scalable, automated systems for end-to-end evaluation of personalized Agentic Search.
  • Design and implement offline and online metrics to assess product and component quality for Personal Question Answering.
  • Develop data generation and curation methods to improve the quality of search results.
  • Create comprehensive metrics frameworks for continuous quality assessment and iterative product improvements.
  • Define the long-term technical vision and roadmap for personalized Siri Search quality.
  • Identify technical gaps and drive solutions to enhance the Personal Question Answering stack.
  • Partner with cross-functional teams to define data requirements and prioritize evaluation goals.

What we're looking for

  • Bachelor's degree in Computer Science, Engineering, Statistics, or a related field.
  • Master's or Ph.D. in Computer Science, Engineering, Statistics, or a related field (preferred).
  • 6 or more years of industry experience building machine learning or machine learning evaluation systems at scale.
  • 2 or more years of industry experience in technical leadership or management.
  • Knowledge of machine learning fundamentals, neural network architectures, and software engineering in Python or C/C++.
  • Communication and collaboration skills.
  • Experience designing and building production machine learning systems in search, NLP, recommendation systems, or information retrieval (preferred).
  • Experience in data collection, generation, or quality assessment of language, image, or multimodal data (preferred).

More like this

Similar roles

Senior Machine Learning Engineering Manager, Evaluation

Apple Inc

Cupertino, CA 34 days ago $237,600–$401,700
Python Large Language Models LLM-as-judge Reinforcement Learning Synthetic Data Generation Agentic Systems Reward Modeling Post-training Multi-objective Optimization Deep Learning Frameworks Benchmark Infrastructure Simulation Environments Trajectory Analysis
8+ yrs exp