Machine Learning Applied Researcher, Speech, Vision and Audio

Apple Inc

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
Cambridge, MA
Salary
$165,800–$249,500 / yr
Posted
8 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $233k
This role $208k
$150k most similar roles pay here $309k

This role pays less than 75% of similar roles. Most pay $211,612–$254,812 — the shaded band above. At the midpoint, this role pays about $208k versus about $233k for comparable roles.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Machine Learning Applied Researcher, Speech, Vision and Audio

As a Machine Learning Applied Researcher - Speech, Vision and Audio, you will join a dynamic team focused on researching new capabilities and developing frontier models using cutting-edge methods. You will tackle high-risk, high-reward challenges by designing and improving state-of-the-art deep learning models trained on unique data while implementing novel machine learning algorithms to solve problems without obvious answers. Your daily work involves navigating ambiguous problem spaces, formulating hypotheses, and iterating toward solutions while balancing performance with deployment constraints like latency and memory. You will work across multiple modalities including vision, audio, sensor, and text data to improve automatic speech recognition. The role requires proficiency in Pytorch and expertise in areas such as self-supervised learning, synthetic data generation, large language model training, and multimodal data processing to solve complex technical challenges.

What you'll do

  • Research and develop cutting-edge deep learning models using vision, audio, sensor, and text data.
  • Design and implement novel neural network architectures specifically for automatic speech recognition (ASR).
  • Formulate hypotheses and run experiments to solve problems in ambiguous research spaces.
  • Optimize model performance while adhering to deployment constraints like latency, memory, and compute.
  • Translate complex research ideas into production-quality code.
  • Identify and integrate frontier machine learning methods from scientific literature into existing products.
  • Synthesize multi-disciplinary approaches to create solutions that can be adopted across different teams.

What we're looking for

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related field (or equivalent practical experience).
  • Master's or PhD in Computer Science, Electrical Engineering, or a related field with 10+ years of relevant industry experience.
  • Experience in academic or industry research.
  • Proficiency and experience working with PyTorch.
  • Strong foundation in deep learning theory and hands-on experience training large-scale models.
  • Deep knowledge in self-supervised learning, synthetic data generation, LLM training, or automatic speech recognition.
  • Experience working with multimodal data including images, audio, time-series, or sensor fusion.
  • Publication record at top-tier ML venues such as NeurIPS, ICML, ICLR, Interspeech, or ICASSP.

More like this

Similar roles

Applied AI Scientist, Multimodal Intelligence

Apple Inc

Sunnyvale, CA 45 days ago $216,200$324,800
Generative AI Multimodal Foundation Models Deep Learning Python PyTorch JAX Computer Vision Data Pipelines On-device Machine Learning Vision-Language Models Video-Language Models
6+ yrs exp

Senior Machine Learning Scientist, Siri Speech

Apple Inc

Cupertino, CA 148 days ago $184,700$324,800
Deep Learning Python PyTorch JAX TensorFlow Natural Language Processing Speech Recognition multi-modal learning Conversational AI Audio Modeling Foundation Models Machine Learning Training

Senior Machine Learning Scientist, Siri Speech

Apple Inc

Cupertino, CA 84 days ago $184,700$324,800
Deep Learning Python PyTorch JAX TensorFlow Natural Language Processing Speech Recognition multi-modal learning Conversational AI Audio Modeling Machine Learning Foundation Models

Senior Applied ML Researcher, Video Apps

Apple Inc

Cupertino, CA 34 days ago $184,700$324,800
Deep Learning Machine Learning Python PyTorch Computer Vision Audio Signal Processing Multimodal Learning Generative AI Neural Networks Objective-C Swift Linear Algebra Probability Optimization
8+ yrs exp