Machine Learning Applied Researcher, Speech, Vision and Audio

Apple Inc

Confirmed live 2 days ago High trust

Quick summary

Work type
On-site
Location
Cambridge, MA
Salary
$134,700–$203,000 / yr
Posted
8 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $233k
This role $169k
$116k most similar roles pay here $313k

This role pays less than 92% of similar roles. Most pay $211,612–$254,812 — the shaded band above. At the midpoint, this role pays about $169k versus about $233k for comparable roles.

Based on 240 similar postings.

Employer

About Apple Inc

Apple Inc. is a multinational technology company known for designing and manufacturing consumer electronics, software, and online services, including the iPhone, Mac, iPad, and App Store. Industry: Consumer Electronics & Software

Apple Inc currently has 1984 open roles on FindRole.

Listed pay typically runs $175,000–$277,600 across 1590 roles with salary data.

Most-posted roles

View all roles at Apple Inc

At a glance

TL;DR · Machine Learning Applied Researcher, Speech, Vision and Audio

Machine Learning Applied Researcher - Speech, Vision and Audio will join a dynamic team to research new capabilities and develop frontier models using cutting-edge methods. The role involves tackling high-risk, high-reward challenges by researching, designing, and improving state-of-the-art deep learning models trained on unique data. Day-to-day responsibilities include implementing novel neural network architectures for automatic speech recognition, navigating ambiguous problem spaces to formulate hypotheses, and optimizing model performance against latency, memory, and compute constraints. The candidate will work with multimodal data including vision, audio, sensor, and text. Required skills include proficiency in PyTorch and experience with large-scale models, self-supervised learning, synthetic data generation, or large language model training. This role focuses on solving complex problems in the field of automatic speech recognition by synthesizing approaches from different fields to create solutions adopted across the organization.

What you'll do

  • Research and develop cutting-edge deep learning models using vision, audio, sensor, and text data.
  • Design and implement novel neural network architectures specifically for automatic speech recognition (ASR).
  • Formulate hypotheses and run experiments to solve problems in ambiguous research spaces.
  • Optimize model performance while accounting for deployment constraints like latency, memory, and compute.
  • Translate complex research ideas into production-quality code.
  • Identify and integrate frontier machine learning methods from current scientific literature into practical solutions.

What we're looking for

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related field (or equivalent practical experience).
  • Master's or PhD in Computer Science, Electrical Engineering, or a related field.
  • Experience in academic or industry research.
  • Proficiency and experience working with PyTorch.
  • Strong foundation in deep learning theory and hands-on experience training large-scale models.
  • Deep knowledge in self-supervised learning, synthetic data generation, LLM training, or automatic speech recognition.
  • Experience working with multimodal data including images, audio, time-series, or sensor fusion.
  • Proven track record of publications at top-tier ML venues like NeurIPS, ICML, ICLR, Interspeech, or ICASSP.

More like this

Similar roles

Applied AI Scientist, Multimodal Intelligence

Apple Inc

Seattle, WA 45 days ago $205,400$308,500
Generative AI Multimodal Foundation Models Deep Learning Python PyTorch JAX Computer Vision Data Pipelines On-device Machine Learning Data Curation
6+ yrs exp

Senior Machine Learning Scientist, Siri Speech

Apple Inc

Cupertino, CA 148 days ago $184,700$324,800
Deep Learning Python PyTorch JAX TensorFlow Natural Language Processing Speech Recognition multi-modal learning Conversational AI Audio Modeling Foundation Models Machine Learning Training

Senior Applied ML Researcher, Video Apps

Apple Inc

Cupertino, CA 34 days ago $184,700$324,800
Deep Learning Machine Learning Python PyTorch Computer Vision Audio Signal Processing Multimodal Learning Generative AI Neural Networks Objective-C Swift Linear Algebra Probability Optimization
8+ yrs exp

Speech & Audio Research Engineer

Qualcomm

San Diego, CA 80 days ago $148,300$222,500
Digital Signal Processing Machine Learning Deep Learning C++ Python TensorFlow PyTorch Audio Processing CNN RNN LSTM GRU Transformer ARM Quantization Linux Unix Signal Modeling Automatic Speech Recognition Text-to-Speech
4+ yrs exp