Senior Deep Learning Scientist, Multimodal Agentic RL
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 66% of similar roles. Most pay $185,900–$254,750 — the blue band above. At the midpoint, this role pays about $236k versus about $220k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 1463 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 1096 roles with salary data.
Most-posted roles
At a glance
The Senior Deep Learning Scientist, Multimodal Agentic RL joins the Nemotron LLM team to advance streaming and agentic multimodal AI. You will develop, train, fine-tune, and deploy large language models for agentic systems capable of audio-visual reasoning, tool usage, and document understanding. Key responsibilities include advancing post-training and alignment methods like instruction tuning, preference optimization, and RLHF/RLVR/MOPD to improve multimodal agents for complex use cases. You will research agentic reasoning, grounded perception, and long-horizon task completion. Required skills include Python, PyTorch, and expertise in Transformers, mixture-of-experts models, and reinforcement learning algorithms like MDPs and reward design. You will work on the Nemotron Omni and VoiceChat platforms to solve real-world challenges in planning and tool execution across digital and physical environments.
Skills
What you'll do
What we're looking for
More like this
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia