Senior Deep Learning Scientist, Multimodal Agentic RL
Nvidia
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 66% of similar roles. Most pay $185,900–$254,750 — the blue band above. At the midpoint, this role pays about $236k versus about $220k for comparable roles.
Based on 240 similar postings.
Employer
Nvidia is a leading designer of graphics processing units (GPUs) and system-on-chip units, powering gaming, professional visualization, data centers, and artificial intelligence workloads. Industry: Semiconductors & AI Computing
Nvidia currently has 1463 open roles on FindRole.
Listed pay typically runs $184,000–$287,500 across 1096 roles with salary data.
Most-posted roles
At a glance
The Senior Deep Learning Scientist, Multimodal Agentic RL joins the Nemotron LLM team to advance streaming and agentic multimodal AI. You will develop, train, fine-tune, and deploy large language models for agentic systems capable of audio-visual reasoning, tool usage, and document understanding. Key responsibilities include advancing post-training and alignment methods like instruction tuning, preference optimization, and RLHF/RLVR/MOPD to improve multimodal agents for complex use cases. You will research agentic reasoning, grounded perception, and long-horizon task completion. Required skills include Python, PyTorch, and expertise in Transformers, mixture-of-experts models, and reinforcement learning algorithms like MDPs and reward design. You will work on the Nemotron Omni and VoiceChat platforms to build models that perform planning and tool execution across digital and physical environments.
Skills
What you'll do
What we're looking for
More like this
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia
Nvidia