Director, AI Research - Recursive Self-Improvement
Amd
Quick summary
Market check
How this pay compares to similar roles
This role pays more than 75% of similar roles. Most pay $187,390–$256,237 — the shaded band above. At the midpoint, this role pays about $255k versus about $222k for comparable roles.
Based on 240 similar postings.
Employer
AMD (Advanced Micro Devices) is a semiconductor company that develops high-performance processors, graphics cards, and adaptive computing solutions for gaming, data centers, and embedded markets. Industry: Semiconductors
Amd currently has 367 open roles on FindRole.
Listed pay typically runs $166,400–$249,600 across 367 roles with salary data.
Most-posted roles
At a glance
As an AI Research Scientist, Recursive Self Improvement, AI Safety and Reinforcement Learning, you will join the team to research recursive self-improvement in a bounded, engineering-first context. You will investigate systems where models, data generators, or toolchains improve their own training signals, curricula, or verification under explicit governance and human oversight. Your daily work involves researching self-improving training loops like model-generated supervision and iterative distillation while designing measurement and containment for RSI pipelines to prevent bias or reward hacking. You will develop evaluations for capability drift and Goodhart effects, partner with RL scientists on policy optimization, and define red-team protocols for monitoring. The role requires expertise in machine learning, AI safety, reinforcement learning, and software skills to build experimental harnesses. This work addresses the technical challenges of ensuring auditable, safe training for hardware and generative-AI programs.
Skills
What you'll do
What we're looking for
More like this
Amd
Amd
Amd
Amd
Anduril Industries
Amd