| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| GEAR (Generalist Embodied Agent Research) | Jim Fan (Linxi Fan) and Yuke Zhu | GR00T N1 (2025); Voyager (2024) | A | ||
| DAIR (Data-Driven AI for Robotics) | Umar Iqbal | GLAMR (2022); GENMO (2025) | A | ||
| RVP (Robotic Visual Perception Research) | Stan Birchfield | DOPE (2018); FoundationPose (2024) | A | ||
| DLER (Deep Learning Efficiency Research) | Pavlo Molchanov | Minitron (2024); Hymba (2024) | A | ||
| GenAIR (Fundamental Generative AI Research) | Arash Vahdat | NVAE (2020); LSGM (2021) | A | ||
| Multimodal / VLM group (no public name, no lab page) | Zhiding Yu | SegFormer (2021); Eagle 2 (2025) | B |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Abhishek Badki | Senior Research Scientist | Low-level vision, 4D perception | L4P (2026); Zero-shot Monocular Scene Flow (2025); Binary TTC (2021) | org · site | A |
| De-An Huang | Research Scientist | Multimodal LLMs, embodied AI | Eagle 2.5 (2025); FRAG (2025) | site | D |
| Guilin Liu | Principal Research Scientist & Team Leader | Multimodal LLMs, VLMs | Partial Conv Inpainting (2018); Eagle 2 (2025); LocateAnything (2026) | site | D |
| Hang Su | Research Scientist | 2D/3D representation learning | SPLATNet (2018); Zero-shot Monocular Scene Flow (2025); L4P (2026) | org · site | A |
| Jan Kautz | Vice President, Learning and Perception Research | Vision, learning, embodied AI | Precomputed Radiance Transfer (2002); PWC-Net (2018); GR00T N1 (2025) | org · site | A |
| Jim Fan (Linxi Fan) | Director of Robotics & Distinguished Research Scientist, GEAR co-l | Embodied foundation models | MineDojo (2022); Voyager (2024); Eureka (2024) | org · site | A |
| Jindong Jiang | Research Scientist | Multimodal LLMs, vision foundation models | STORM (2025); Slot State Space Models (2024); Object-Centric Slot Diffusion (2023) | org · site | A |
| Jinwei Gu | Principal Research Scientist & Senior Manager (now NVIDIA Cosmos L | Computational imaging, world models | NVIDIA Cosmos (2025); Spatial Propagation Networks (2017) | org · site | A |
| Shalini De Mello | Director of Research, New Experiences | Human-computer interaction, 4D worlds | EG3D (2022); GroupViT (2022); Few-Shot Adaptive Gaze Estimation (2019) | org | A |
| Shizhe Diao | Research Scientist | Data quality and data efficiency | Hymba (2024); ProRL (2025) | site | A |
| Sifei Liu | Principal Research Scientist & Tech Lead | Spatial reasoning, embodied foundation models | GroupViT (2022); SCOPS (2019); SpatialRGPT (2024) | org · site | A |
| Xin Dong | Research Scientist | On-device language models | Hymba (2024); Small Language Models are the Future of Agentic AI (2025) | — | A |
| Zhiding Yu | Principal Research Scientist & Research Lead | VLM/VLA, multimodal LLMs | SegFormer (2021); VoxFormer (2023); Eagle (2024) | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Multimodal / VLM thrust (VILA, NVILA, VILA-HD, NaVILA, OmniVinci, Nemotron-Omni) | Hongxu (Danny) Yin | VILA (2024); NVILA (2025) | A | ||
| Efficient LLMs / compression & NAS (Minitron, Flextron, Nemotron-Elastic, Puzzle collaboration) | Saurav Muralidharan | Minitron (2024); Flextron (2024) | D | ||
| Vision foundation models (RADIO / AM-RADIO / RADIOv2.5 / FeatSharp / PHI-S) | Greg Heinrich | AM-RADIO (2024); RADIOv2.5 (2025) | D | ||
| Agentic-system reliability & efficiency (Small Language Models are the Future of Agentic AI, Minifinetuning) | Peter Belcak | Small Language Models are the Future of Agentic AI (2025) | D |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Ali Hatamizadeh | LLM Tech Lead & Staff Research Scientist, NVIDIA Research | LLM architecture design, Gated DeltaNet | Gated DeltaNet (2025); FasterViT (2024); Global Context ViT (2023) | org · site | D |
| Greg Heinrich | Vision foundation and vision-language models | AM-RADIO (2024); FasterViT (2024) | org · site | A | |
| Hanrong Ye | Research Scientist, NVIDIA Research | Omni-modal LLMs, OmniVinci, Nemotron-Omni | OmniVinci (2025); MM-Ego (2025); InvPT (2022) | org · site | A |
| Hongxu (Danny) Yin | Principal Research Scientist & Research Lead, NVIDIA Research | Multimodal LLMs, VILA family, post-training | VILA (2024); NVILA (2025); DeepInversion (2020) | org · site | A |
| Matthijs Van Keirsbilck | Senior Research Scientist, NVIDIA | Efficient architectures, sparsity, quantization | Hymba (2024); Rethinking Full Connectivity in RNNs (2019) | org · site | E |
| Pavlo Molchanov | Research Director / Director of Research, NVIDIA Research (team le | Efficiency of LLMs and multimodal models | Minitron (2024); Hymba (2024); FasterViT (2024) | org · site | A |
| Peter Belcak | Research Scientist, NVIDIA Research | Reliable, efficient agentic systems; SLMs | Small Language Models are the Future of Agentic AI (2025) | org · site | A |
| Saurav Muralidharan | Senior Research Scientist, NVIDIA Research | LLM pruning, distillation, NAS | Minitron (2024); Flextron (2024); MaskLLM (2024) | org · site | A |
| Wonmin Byeon | Sequence models, efficient architectures | GroupViT (2022); STORM (2025); Convolutional Tensor-Train LSTM (2020) | org | E | |
| Yingyan (Celine) Lin | Visiting Professor at NVIDIA (Associate Professor, Georgia Institu | Efficient model/hardware co-design | EyeCoD (2022); Fusion-3D (2024); Omni-Recon (2024) | site | A |
| Yonggan Fu | Senior Research Scientist, NVIDIA Research | Efficient LM architectures, Hymba, Nemotron-Flash | Hymba (2024); Nemotron-Flash (2025); Omni-Recon (2024) | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Generative AI for Proteins and Molecules | Proteina (2025); La-Proteina (2026) | A | |||
| Generative AI for Weather | Elucidated Rolling Diffusion Models (2025); ATLAS (2026) | A | |||
| Compositionality and Control | ProtComposer (2025); DisCo-Diff (2024) | A | |||
| Fundamental Research | LSGM (2021); Denoising Diffusion GANs (2022) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Arash Vahdat | Research Director; GenAIR Team Lead | generative learning, diffusion, gen AI for science | NVAE (2020); LSGM (2021); Denoising Diffusion GANs (2022) | org · site | A |
| Chao Liu | Research Scientist | video generation, diffusion distillation | Live 3D Portrait (2023); BlobGEN-3D (2024) | org | A |
| Julius Berner | Research Scientist | probabilistic ML, sampling, neural PDE solvers | org · site | A | |
| Karsten Kreis | Principal Research Scientist | diffusion models, generative AI for science | Align your Latents (2023); LION (2022); Proteina (2025) | org · site | A |
| Kieran Didi | Research Scientist | protein/binder generative design, ML for chemistry | Proteina (2025); Proteina-Complexa (2026) | org · site | A |
| Morteza Mardani | Principal Scientist (also Visiting Researcher, Stanford University | diffusion/flow models, weather and science AI | ATLAS (2026); Elucidated Rolling Diffusion Models (2025); Warped Diffusion (2024) | org · site | A |
| Tomas Geffner | Research Scientist | probabilistic ML, generative modeling, sampling | Proteina (2025); La-Proteina (2026); Compositional Score Modeling (2023) | org · site | A |
| Yongxin Chen | Research Scientist (also Associate Professor, Georgia Institute of | optimal transport, control, sampling, ML | DEIS (2023); gDDIM (2023); Path Integral Sampler (2022) | org · site | A |
| Zuobai Zhang | Research Scientist | protein foundation models, representation learning | Proteina (2025); Proteina-Complexa (2026) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Davis Rempe | Senior Research Scientist | human and humanoid motion, 3D perception | Trace and Pace (2023); Kimodo (2026) | site | A |
| Haotian Zhang | Senior Research Scientist | 3D human motion perception and generation | Vid2Player (2021); Physically Simulated Tennis Skills (2023); GENMO (2025) | site | A |
| Jiefeng Li | Research Scientist | computer vision, human pose, generative AI | HybrIK (2021); Residual Log-likelihood Estimation (2021); BLADE (2025) | org · site | A |
| Jinhyung (David) Park | Research Scientist | 3D scene understanding, human modeling | GRAIL (2026); SONIC (2026) | org · site | A |
| Kevin Xie | Research Scientist (concurrently PhD student, University of Toront | generative models, 3D vision, humanoid motion | VideoPanda (2025); OmniDreams (2026) | site | A |
| Mathis Petrovich | Research Scientist | text-to-3D human motion synthesis | ACTOR (2021); TEMOS (2022); TMR (2023) | site | A |
| Tianyi Xie | Research Scientist | computer graphics, generative AI, 3D vision | PhysGaussian (2024); GRAIL (2026); PhysAnimator (2025) | org · site | A |
| Umar Iqbal | Sr. Research Manager; DAIR Team Lead | robot learning from human data | GLAMR (2022); GAvatar (2024); GENMO (2025) | org · site | A |
| Xue Bin (Jason) Peng | Research Scientist at NVIDIA AND Assistant Professor, Simon Fraser | character animation, RL, robotics | DeepMimic (2018); AMP (2021); ASE (2022) | site | A |
| Xueting Li | Research Scientist | computer vision, 3D vision | GAvatar (2024); Dream-in-4D (2024); Pyramid Diffusion (2024) | org · site | A |
| Ye Yuan | Senior Research Scientist | 3D vision, embodied AI, reinforcement learning | GLAMR (2022); PhysDiff (2023); SONIC (2026) | org · site | A |
| Yifeng Jiang | Research Scientist | character animation, physics simulation, humanoids | ProtoMotions3 (2025); MaskedManipulator (2025) | site | A |
| Yufei (Judy) Ye | 3D vision, robotics, human motion | Affordance Diffusion (2023); MUSIC (2026) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Bowen Wen | Staff Research Scientist | 3D visual perception for manipulation | FoundationPose (2024); FoundationStereo (2025); BundleSDF (2023) | org · site | D |
| Jeff Smith | Research Engineer | perception/robotics tooling and systems | HANDAL (2023) | org | E |
| Jonathan Tremblay | Research Scientist | synthetic data, pose estimation, robot vision | DOPE (2018); NViSII (2021); Falling Things (2018) | org · site | D |
| Stan Birchfield | Principal Research Scientist and Senior Research Manager; RVP Team | computer vision and robotics intersection | FoundationPose (2024); FoundationStereo (2025); BundleSDF (2023) | org · site | B |
| Stephen Tyree | Research Scientist | 6-DoF pose estimation, benchmarks, robot vision | GA3C (2017); NViSII (2021); BundleSDF (2023) | org · site | D |
| Valts Blukis | Senior Research Scientist | language-grounded robot perception, 3D scenes | ProgPrompt (2023); RVT (2023); CuRobo (2023) | org · site | D |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| World Model team (Project GR00T) | Joel Jang | DreamGen (2025); DreamDojo (2026) | A | ||
| GR00T VLA foundation model / post-training | (no named sub-lead found; Johan Bjorck is repeat first author, Scott Reed was founding senior scientist until 2026) | GR00T N1 (2025); GR00T N1.5 (2025) | D | ||
| Humanoid whole-body control / locomotion | (no named lead found) | SONIC (2026); HOVER (2025) | D | ||
| Foundation agents for games / virtual worlds | (no named lead found; Guanzhi Wang is the through-line) | MineDojo (2022); Voyager (2024) | D | ||
| Embodied reasoning / generalist planner | (no named lead found) | Vesta (2026) | D | ||
| Simulation & synthetic data infrastructure | (no named lead found) | HumanoidMimicGen (2026) | C |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Avnish Narayan | (title not found) | robot learning, RL | GR00T N1 (2025) | org | D |
| Fengyuan Hu | Research Engineer | robot learning infrastructure, VLA training | GR00T N1 (2025); NitroGen (2026) | org | C |
| Fernando Castaneda | Research Scientist (UC Berkeley PhD in control) | humanoid control, sim-to-real | GR00T N1 (2025); SONIC (2026) | org | D |
| Guanzhi Wang | Research Scientist | robot foundation model post-training, gaming agents | Voyager (2024); MineDojo (2022); NitroGen (2026) | org · site | C |
| Jing Wang | (title not found) | VLA models, robot learning | GR00T N1 (2025); Vesta (2026) | org | D |
| Joel Jang | Senior Research Scientist; leads the GEAR world model team for Pro | video world models, robot learning | DreamGen (2025); LAPA (2025); DreamDojo (2026) | site | A |
| Johan Bjorck | Senior Research Scientist (per LinkedIn) | VLA foundation models, pretraining | GR00T N1 (2025); Vesta (2026) | org | D |
| Kaiyuan Zheng | (title not found) | video world models | — | E | |
| Kaushil Kundalia | (title not found) | training infrastructure, VLA | GR00T N1 (2025) | — | D |
| Kevin Lin | Researcher (PhD UT Austin, advised by Yuke Zhu) | manipulation, data generation | GR00T N1 (2025); CP-Gen (2025) | site | C |
| Linxi "Jim" Fan | Co-lead of GEAR; Director of AI & Distinguished Scientist (per Lin | generalist embodied agents, foundation agents | MineDojo (2022); Voyager (2024); Eureka (2024) | org · site | A |
| Loic (Loic) Magne | (title not found) | gaming agents, multimodal models | NitroGen (2026); GR00T N1 (2025) | org | D |
| Mengda Xu | Research Scientist | robotic manipulation, human-to-robot transfer | DexUMI (2025); XSkill (2023); Flow as the Cross-Domain Manipulation Interface (2024) | org · site | C |
| Nikita Cherniadev | Senior Simulation Engineer | simulation infrastructure for embodied AI | GR00T N1 (2025); HumanoidMimicGen (2026) | org · site | C |
| Qi Wang | (title not found) | VLA training, robot data | GR00T N1 (2025) | — | D |
| Ruijie Zheng | (ambiguous - see evidence) | human-video pretraining, VLA scaling | GR00T N1 (2025) | org · site | C |
| Runyu Ding | (ambiguous - see evidence) | 3D scene understanding, robot data | GR00T N1 (2025); HumanoidMimicGen (2026); Bunny-VisionPro (2024) | org · site | D |
| Scott Reed | FORMER - Principal Research Scientist and founding member of GEAR, | generalist agents, VLA models | GR00T N1 (2025) | org · site | B |
| Spencer Huang | (title not found) | robot learning, data pipelines | GR00T N1 (2025) | — | D |
| Xiaowei Jiang | (title not found) | robot learning systems | org | E | |
| Xingye Da | (title not found; consistently NVIDIA-affiliated) | humanoid locomotion, whole-body control | GR00T N1 (2025); SONIC (2026) | — | D |
| Yinzhen Xu | Research Engineer (per LinkedIn) | VLA training, 3D vision | GR00T N1 (2025); NitroGen (2026) | org | D |
| You Liang Tan | Researcher / full-stack roboticist | imitation learning, robot systems | GR00T N1 (2025) | org · site | C |
| Yu Fang | (title not found) | simulation, graphics for robotics | GR00T N1 (2025); HumanoidMimicGen (2026) | org | D |
| Yuke Zhu | Co-lead of GEAR; Director and Distinguished Research Scientist, NV | robot learning, manipulation, embodied AI | MineDojo (2022); Eureka (2024); GR00T N1 (2025) | org · site | A |
| Yunze Man | Research Scientist | VLM post-training, embodied AI | Lexicon3D (2024); SceneCraft (2024); Vesta (2026) | org · site | C |
| Yuqi Xie | (ambiguous - see evidence) | robot manipulation, agents | GR00T N1 (2025); HARMON (2024) | org · site | C |
| Zhe Zhang | (title not found) | VLA training, evaluation | — | E | |
| Zhengyi (Zen) Luo | Senior Research Scientist | humanoid motion, whole-body control | SONIC (2026); HOVER (2025); ASAP (2025) | org · site | C |
| Zhiqi Li | Research Scientist, NVIDIA Research (Singapore) | vision-language models, 3D perception | BEVFormer (2022); InternImage (2023); Eagle 2 (2025) | org · site | C |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| World Simulation | Zian Wang and Jun Gao (world modeling); Jose M. Alvarez, Seung Wook Kim, Xiuming Zhang (AV agents); Xue Bin (Jason) Peng, Yifeng Jiang, Davis Rempe (humanoid agents) | GEN3C (2025); Cosmos-Drive-Dreams (2025) | A | ||
| Physics Simulation | Ken Museth | OpenVDB (2013); NanoVDB (2021) | A | ||
| 4D Perception | Laura Leal-Taixe (with Aljosa Osep, Tim Meinhardt, Guillem Braso, Sergio Agostinho, Qunjie Zhou) | Better Call SAL (2024); Light3R-SfM (2025) | A | ||
| 3D/4D Content Creation & Editing | Zan Gojcic (with Jun Gao, Or Litany, Jonathan Lorraine, Seung Wook Kim, Xuanchi Ren, Frank Shen, Masha Shugrina) | GET3D (2022); LATTE3D (2024) | A | ||
| Geometry Processing | Nicholas Sharp (with Mikaela Angelina Uy, Donglai Xiang, Sangeetha Grama Srinivasan, Tianchang Shen) | FlexiCubes (2023); SpaceMesh (2024) | A | ||
| Human Motion Modeling | Davis Rempe (with Xue Bin (Jason) Peng, Haotian Zhang, Yifeng Jiang, Mathis Petrovich) | ASE (2022); MaskedMimic (2024) | A | ||
| NuRec (Neural Reconstruction Engine) | Zan Gojcic, Kangxue Yin | 3D Gaussian Ray Tracing (2024); 3DGUT (2025) | A | ||
| Spatial Data SDKs (ViPE, NCore) | Jiahui Huang (ViPE); Janick Martinez Esturo (NCore) | ViPE (2025) | A | ||
| Spatial AI SDKs (Kaolin, fVDB) | Clement Fuji Tsang and Masha Shugrina (Kaolin); Francis Williams (fVDB) | fVDB (2024) | A | ||
| Sparse Volume SDKs (OpenVDB, NanoVDB, fVDB) | Ken Museth, Francis Williams | OpenVDB (2013); NanoVDB (2021) | A | ||
| Creative and Applied AI Tools (CAAT) | Masha Shugrina | ATISS (2021); Simplicits (2024) | A | ||
| Autonomous Vehicle Applied Research (AVAR) | Jose M. Alvarez | SegFormer (2021); OmniDrive (2025) | A | ||
| High-Fidelity Physics Research Group (prl) | Ken Museth | OpenVDB (2013); NanoVDB (2021) | A | ||
| Dynamic Vision and Learning Group (dvl) | Laura Leal-Taixe | MOTChallenge (2015); TrackFormer (2022) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Amirmojtaba Sabour | 3D generation, data curation | VideoPanda (2025) | — | A | |
| Amlan Kar | Senior Research Scientist | data-centric AI, autolabeling | Meta-Sim (2019); Neural Turtle Graphics (2019); Meta-Sim2 (2020) | site | A |
| Anita Hu | 3D deep learning | Painting with 3D Gaussian Splat Brushes (2025) | — | A | |
| Clement Fuji Tsang | Senior Research Engineer (Kaolin) | Kaolin 3D deep learning library | ArtisanGS (2026) | — | A |
| Davis Rempe | Senior Research Scientist | human/humanoid motion generation | Trace and Pace (2023); GENMO (2025); Kimodo (2026) | site | A |
| Despoina Paschalidou | Senior Research Scientist | 3D scene generation | Superquadrics Revisited (2019); Neural Parts (2021); ATISS (2021) | site | A |
| Donglai Xiang | Research Scientist | digital humans, clothed avatars | Monocular Total Capture (2019); Dressing Avatars (2022); PartField (2025) | site | A |
| Erwin Coumans | physics simulation, robotics | — | A | ||
| Haithem Turki | Senior Research Scientist | large-scale radiance fields | Mega-NeRF (2022); SUDS (2023); HybridNeRF (2024) | org · site | A |
| Haotian Zhang | Senior Research Scientist | human motion, humanoid control | Vid2Player (2021); Learning Physically Simulated Tennis Skills (2023); GENMO (2025) | org · site | A |
| Hassan Abu Alhaija | Machine Learning Researcher / Research Engineer | synthetic media, generative 3D | Enhancing Photorealism Enhancement (2023); Augmented Reality Meets Computer Vision (2018) | site | A |
| Janick Martinez Esturo | Senior Research Engineer (Munich) | NCore multi-sensor data platform | 3D Gaussian Ray Tracing (2024); 3DGUT (2025) | site | A |
| Jiahui Huang | Senior Research Scientist | 3D scene reconstruction, ViPE | Neural Kernel Surface Reconstruction (2023); XCube (2024); ViPE (2025) | site | A |
| Jialiang Wang | 3D vision, depth estimation | Emu (2023); Movie Gen (2024) | org · site | A | |
| Jiawei Ren | Research Scientist | 4D generation, dynamic scenes | DreamGaussian (2024); L4GM (2024); BTimer (2025) | site | A |
| Jonathan Lorraine | Research Scientist | fast text-to-3D, LATTE3D | ATT3D (2023); LATTE3D (2024) | site | A |
| Jun Gao | Research Scientist | 3D generative models, GET3D | DMTet (2021); GET3D (2022); FlexiCubes (2023) | site | A |
| Kangxue Yin | Senior Research Scientist | neural reconstruction, 3D generation | DMTet (2021); 3DStyleNet (2021); GET3D (2022) | site | A |
| Katarina Tothova | 3D vision | DiffusionHarmonizer (2026) | — | A | |
| Kevin Xie | Research Scientist | generative world models, OmniDreams | ATT3D (2023); LATTE3D (2024); VideoPanda (2025) | site | A |
| Masha Shugrina | Researcher / group lead, Creative and Applied AI Tools (CAAT) | creative AI tools, Kaolin | ATISS (2021); Painting with 3D Gaussian Splat Brushes (2025) | site | A |
| Matan Atzmon | Research Scientist | neural implicit shape representations | SAL (2020); Implicit Geometric Regularization (2020); Frame Averaging (2022) | site | A |
| Mathis Petrovich | Research Scientist (Zurich) | text-to-3D human motion | ACTOR (2021); TEMOS (2022); TMR (2023) | site | A |
| Michal Tyszkiewicz | 3D vision, feature matching | TokenGS (2026) | site | A | |
| Mikaela Angelina Uy | Research Scientist | 3D shape analysis, reconstruction | ScanObjectNN (2019); Point2Cyl (2022); PartField (2025) | site | A |
| Nicholas Sharp | Senior Research Scientist (Seattle) | GPU geometry processing | DiffusionNet (2022); FlexiCubes (2023); Adaptive Shells (2023) | site | A |
| Or Litany | Senior Research Scientist | 3D vision, generative AI | VoteNet (2019); PointContrast (2020); LION (2022) | site | A |
| Or Perel | Researcher | neural 3D representations | 3D Gaussian Ray Tracing (2024); Simplicits (2024); Painting with 3D Gaussian Splat Brushes (2025) | site | A |
| Qi Wu | Senior Research Scientist | large-scale neural rendering | 3DGUT (2025); SimULi (2026) | site | A |
| Riccardo de Lutio | Senior Research Scientist | 3D vision, neural reconstruction | Guided Super-Resolution (2019); 3D Gaussian Ray Tracing (2024); OmniRe (2025) | site | A |
| Ruilong Li | Research Scientist | radiance fields, 3D reconstruction | AIST++ (2021); Nerfstudio (2023); gsplat (2025) | site | A |
| Sangeetha Grama Srinivasan | differentiable simulation, digital humans | Learning Active Quasistatic Physics-Based Models (2021); Deep 3D Simulation Super-Resolution (2024) | site | A | |
| Sanja Fidler | VP of AI Research, NVIDIA; Associate Professor, University of Toro | lab lead; spatial intelligence | GET3D (2022); Meta-Sim (2019); MovieQA (2016) | site | A |
| Seung Wook Kim | Research Scientist | world models, generative simulation | NeuralField-LDM (2023); L4GM (2024) | site | A |
| Tianchang (Frank) Shen | Research Scientist | mesh extraction, FlexiCubes | DMTet (2021); FlexiCubes (2023); Adaptive Shells (2023) | site | A |
| Tianshi Cao | 3D generative modeling | TexFusion (2023); LATTE3D (2024) | — | A | |
| Vismay Modi | physics-based simulation, geometry | EMU (2021); Simplicits (2024) | site | A | |
| Xinglong (Alex) Sun | AI Researcher (San Jose) | AV perception, efficient models | Multi-Dimensional Pruning (2024); AllTracker (2025) | site | A |
| Xiuming Zhang | Principal Research Scientist | world models, RL, vision | MoSculp (2018); NeRFactor (2021); DiffusionRig (2023) | site | A |
| Xuanchi Ren | Senior Research Scientist | 3D-aware video generation, GEN3C | XCube (2024); GEN3C (2025); Lyra (2026) | site | A |
| Xue Bin (Jason) Peng | Research Scientist; Assistant Professor, Simon Fraser University | physics-based humanoid control | DeepMimic (2018); AMP (2021); ASE (2022) | site | A |
| Yifan Lu | Researcher | AV data curation, simulation | InfiniCube (2025); Cosmos-Drive-Dreams (2025); ChronoEdit (2026) | site | A |
| Yifeng Jiang | Research Scientist | physics-based character control | Transformer Inertial Poser (2022); DROP (2023); ProtoMotions3 (2025) | site | A |
| Zan Gojcic | Director of Research (Zurich) | neural reconstruction, world sim | PREDATOR (2021); 3D Gaussian Ray Tracing (2024); 3DGUT (2025) | site | A |
| Zhangjie Wu | Senior Research Scientist | video diffusion, world models | Tune-A-Video (2023); Show-1 (2024); Difix3D+ (2025) | site | A |
| Zian Wang | Senior Research Scientist and Research Manager | inverse rendering, world models | FEGR (2023); Adaptive Shells (2023); DiffusionRenderer (2025) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Physics Simulation (its role inside SIL) | Ken Museth | Monte Carlo Geometry Processing (2020); Walk on Stars (2023) | A | ||
| Sparse Volume SDKs (OpenVDB, NanoVDB, fVDB) | Ken Museth, Francis Williams | OpenVDB (2013); NanoVDB (2021) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Andre Pradhana | Research Scientist | MPM, differentiable simulation, OpenVDB | Drucker-Prager Sand Animation (2016); Multi-Species Sand and Water Mixtures (2017); GPU Optimization of Material Point Methods (2018) | org · site | A |
| Andrew Reidmeyer | Engineer | sparse voxel sim, PNanoVDB, NVIDIA Flow | org | A | |
| Christopher Horvath | Research Scientist / simulation | fluid and water simulation | org | A | |
| David I.W. Levin | Principal Research Scientist, NVIDIA; Associate Professor, Univers | numerical simulation, Simplicits | Simplicits (2024); Neurally Integrated Finite Elements (2025) | org · site | A |
| Denis Zorin | Visiting Faculty; Silver Professor of CS and Mathematics, NYU Cour | geometric modeling, scientific computing | Interactive Multiresolution Mesh Editing (1997); TetWild (2018); Incremental Potential Contact (2020) | org · site | A |
| Ed Quigley | Engineer | physically-based simulation for VFX | Codimensional Surface Tension Flow (2014); Real-Time Interactive Tree Animation (2017) | org · site | A |
| Eftychios Sifakis | Professor of Computer Sciences, University of Wisconsin-Madison (N | FEM, digital humans, solvers | Efficient Elasticity for Character Skinning (2011); SPGrid (2014); FEM Simulation of 3D Deformable Solids (2015) | org · site | A |
| Francis Williams | Senior Research Scientist (NYC) | fVDB, sparse 3D deep learning | Neural Kernel Surface Reconstruction (2023); XCube (2024); fVDB (2024) | org · site | A |
| Gergely (Greg) Klár | Research Scientist | MPM, ML for animation | Drucker-Prager Sand Animation (2016); Shape Targeting (2020); fVDB (2024) | org · site | A |
| Gilles Daviet | Research Scientist | hair/granular sim, differentiable solvers, Warp | Semi-Implicit MPM for Granular Materials (2016); Simple and Scalable Frictional Contacts (2020); Loki (2022) | org · site | A |
| Jonathan Leaf | Research Scientist | cloth simulation, differentiable physics | Interactive Design of Periodic Yarn-Level Cloth Patterns (2018); Weavecraft (2020) | org · site | A |
| Jonathan Swartz | Engineer (ML + simulation) | ML/simulation intersection, OpenVDB | fVDB (2024); Deep 3D Simulation Super-Resolution (2024) | org · site | A |
| Ken Museth | Team Leader; Sr. Director, High-Fidelity Physics Research | OpenVDB/NanoVDB, sparse volumes | VDB (2013); OpenVDB (2013); NanoVDB (2021) | org · site | A |
| Mark Harris | Distinguished Engineer | GPU-accelerated simulation, CUDA | GPGPU Survey (2007); Optimizing Parallel Reduction in CUDA (2007); Parallel Prefix Sum (Scan) with CUDA (2007) | org · site | A |
| Matthew Cong | Senior Research Scientist | facial animation, muscle/flesh simulation | Fully Automatic Anatomical Face Simulation Models (2015); Art-Directed Muscle Simulation (2016); fVDB (2024) | org · site | A |
| Petra Hapalova | Engineer | Omniverse Flow, point cloud tooling | org | A | |
| Rohan Sawhney | Senior Research Scientist | Monte Carlo geometry processing, PDEs | Boundary First Flattening (2018); Monte Carlo Geometry Processing (2020); Walk on Stars (2023) | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| 4D Perception (its role inside SIL) | Laura Leal-Taixé | MOTChallenge (2015); TrackFormer (2022) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alessandro Burzio | PhD Student (joint, Università di Modena e Reggio Emilia) | 4D dynamic scene reconstruction | Déjà View (2026) | org | A |
| Aljosa Osep | Senior Research Scientist | open-world 4D scene understanding | 4D Generic Video Object Proposals (2020); Opening up Open-World Tracking (2022); Better Call SAL (2024) | org · site | A |
| Benjamin Missaoui | PhD Student | video understanding, object tracking | SAMCLR (2023) | org · site | A |
| Cristiano Saltori | Research Scientist | LiDAR domain adaptation, 3D perception | SF-UDA3D (2020); CoSMix (2022); GIPSO (2022) | org · site | A |
| Guillem Brasó | Research Scientist | graph-based multi-object tracking | Learning a Neural Solver for Multi-Object Tracking (2020); MOTSynth (2021); PolarMOT (2022) | org · site | A |
| Laura Leal-Taixé | Team Leader; Senior Research Manager, NVIDIA; Adjunct Professor, T | dynamic scene understanding, tracking | MOTChallenge (2015); HOTA (2020); TrackFormer (2022) | org · site | A |
| Qunjie Zhou | Research Scientist | visual localization, feature matching | Patch2Pix (2021); ViPE (2025) | org · site | A |
| Rodrigo Marcuzzi | Postdoctoral Researcher | LiDAR panoptic segmentation | Mask-Based Panoptic LiDAR Segmentation (2023); Mask4D (2023) | org · site | A |
| Shengyu Huang | Research Scientist | LiDAR/3D scene flow, reconstruction | PREDATOR (2021); Neural LiDAR Fields (2023) | org · site | A |
| Sven Elflein | PhD Student (joint, University of Toronto) | 3D computer vision, feed-forward SfM | MG-GAN (2021); VGG-T3 (2026) | org · site | A |
| Sérgio Agostinho | Senior Research Engineer | pose estimation, geometric deep learning | VGG-T3 (2026) | org · site | A |
| Tim Meinhardt | Research Scientist | multi-object tracking, video segmentation | TrackFormer (2022); Better Call SAL (2024) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alex (Xinglong) Sun | efficient perception models | Multi-Dimensional Pruning (2024); AllTracker (2025) | site | A | |
| Jenny Schmalfuss | robustness, optical flow | Spring (2023); Distracting Downpour (2023); RobustSpring (2026) | site | A | |
| Jose M. Alvarez | Director of AV Applied Research | AV perception, efficient models, data | SegFormer (2021); OmniDrive (2025); GTRS (2025) | site | A |
| Joshua Chen | AV perception | GTRS (2025) | site | A | |
| Maying Shen | model efficiency, pruning | Hardware-Aware Latency Pruning (2023) | site | A | |
| Nadine Chang | long-tail perception, data-centric AV | DriveCritic (2026); GHOST (2026) | site | A | |
| Rafid Mahmood | data acquisition, optimization for AV | Optimizing Data Collection for Machine Learning (2025); AutoScale (2025) | site | A | |
| Shiyi Lan | 3D detection, BEV perception | FastMask (2017); SaccadeNet (2020); DiscoBox (2021) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Generator (world/video generation) | Yogesh Balaji (research manager), Ting-Chun Wang (principal RS) | Cosmos-Predict1 (2025); Cosmos-Predict2.5 (2025) | C | ||
| Reasoner (Cosmos-Reason VLM) | Tsung-Yi Lin | Cosmos-Reason1 (2025); Cosmos-Reason2 (2026) | E | ||
| 3D, camera and geometry | Chen-Hsuan Lin | Neuralangelo (2023); Magic3D (2023) | E | ||
| World models for Physical AI (policy / robotics transfer) | Jinwei Gu | Cosmos (2025); Cosmos 3 (2026) | E | ||
| Transfer / controllable generation | Trung Pham | Cosmos-Transfer1 (2025); Cosmos-Transfer2.5 (2025) | E | ||
| Data curation, embeddings and infrastructure | Francesco Ferroni (with Jing Zhang on training/serving infra) | E | |||
| Audio / omnimodal | Siddharth Gururani | Cosmos 3 (2026) | E |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Akash Gokul | Research Scientist | Generator SFT data | org | C | |
| Alice Luo | Research Scientist | Data pipeline, synthetic data | org | B | |
| Andrew Z. Wang | Research Scientist | Captioning, prompt upsampling | org | B | |
| Chen-Hsuan Lin | Research Scientist & Manager | 3D reconstruction, neural rendering | Neuralangelo (2023); Magic3D (2023); BARF (2021) | org · site | A |
| Chia-Wen Kuo | Research Scientist | Multimodal pretraining data | org | C | |
| David W. Romero | Research Scientist | Efficient generative models | Meshtron (2024) | — | B |
| Fitsum Reda | Principal Scientist | Video tokenization, generative models | FILM (2022); TryOnDiffusion (2023); SDC-Net (2018) | site | B |
| Francesco Ferroni | Principal Researcher | Video-text embeddings, data pipeline | org · site | B | |
| Hamid Eghbalzadeh | Research Scientist | Prompt upsampling, evaluation | org | C | |
| Haotian Zhang | Senior Research Scientist | Vision-language grounding | org | A | |
| Imad El Hanafi | Research Scientist | Large-scale training/inference systems | org | B | |
| Jacob Huffman | Research Scientist | Data infrastructure | org | B | |
| Jiaojiao Fan | Research Scientist | Controllable video generation | RefDrop (2024); Edify Image (2024) | org · site | A |
| Jiashu Xu | Research Scientist | Video captioning, data | org | B | |
| Jiaxiang Tang | Research Scientist | 3D reconstruction and generation | DreamGaussian (2024); LGM (2024); PartPacker (2025) | org · site | A |
| Jing Zhang | Research Scientist / Engineering lead | Training and data infrastructure | Cosmos (2025) | org | B |
| Jinwei Gu | Principal Research Scientist & Senior Manager | World models for Physical AI | Cosmos (2025); Cosmos 3 (2026); Cosmos Policy (2026) | org · site | A |
| Kuno Kim | Research Scientist | Reasoner pretraining, data | org | C | |
| Mengyao Xu | Research Scientist | Captioning, data curation | NV-Embed (2024) | org | C |
| Ming-Yu Liu | Vice President of Research; leads Cosmos Lab | World foundation models, generative AI | GauGAN/SPADE (2019); Video-to-Video Synthesis (2018); Cosmos (2025) | org · site | A |
| Prithvijit Chattopadhyay | Research Scientist | Synthetic data, robust vision | Cosmos-Predict2.5 (2025); SkyScenes (2024); LANCE (2023) | org · site | A |
| Qianli Ma | Senior Research Scientist & Tech Lead | World model pretraining, distillation | CAPE (2020); SCALE (2021); POP (2021) | org · site | A |
| Qinsheng Zhang | Senior Research Scientist | Diffusion sampling, distillation | DEIS (2023); Cosmos (2025) | site | A |
| Sameer Dharur | Research Scientist | VLMs that reason about the world | org | A | |
| Sergei Vasilev | Engineer / Research | Data infrastructure | org | C | |
| Seungjun Nah | Senior Research Scientist | Image/video generation quality | eDiff-I (2022); REDS dataset (2019) | org · site | A |
| Siddharth Gururani | Senior Research Scientist | Audio-visual generative models | Cosmos 3 (2026); SPACE (2023); Fugatto (2025) | org · site | A |
| Stella Shi | Research Scientist | Image/video data curation | org | B | |
| Ting-Chun Wang | Principal Research Scientist | Video and image synthesis | pix2pixHD (2018); Video-to-Video Synthesis (2018); GauGAN/SPADE (2019) | org · site | A |
| Trung Pham | Principal Research Scientist | Controllable world generation | NVAutoNet (2024); Cosmos 3 (2026) | site | A |
| Tsung-Yi Lin | Principal Research Scientist | Physical-AI reasoning VLMs | COCO (2014); Focal Loss (2017); Feature Pyramid Networks (2017) | org · site | B |
| Xiaodong Yang | Principal Research Scientist | Driving world models, action data | MoCoGAN (2018); STEP (2019) | org | B |
| Xiaohui Zeng | Research Scientist | VLM/LLM, multimodal generation | LION (2022); Magic3D (2023) | site | A |
| Xin Kong | Research Scientist | Foundation models for robotics | EscherNet (2024); vMAP (2023) | org · site | A |
| Xingqian Xu | Research Scientist | Image generation, post-training | org | C | |
| Xuan Li | Research Scientist | Physics-based 3D generation | PAC-NeRF (2023); PhysGaussian (2024) | org · site | A |
| Yatian Pang | Research Scientist | Data curation, serving | org | C | |
| Yifan Ding | Research Scientist | Data curation, evaluation | org | B | |
| Yin Cui | Research Scientist | Multimodal data, generative AI | Class-Balanced Loss (2019); iNaturalist Dataset (2018); Simple Copy-Paste (2021) | org · site | B |
| Yogesh Balaji | Staff Research Scientist & Research Manager | Large-scale generative pretraining | eDiff-I (2022); Cosmos (2025) | org · site | A |
| Yu Wang | Research Scientist | Grounding, reasoning evaluation | org | C | |
| Yu-Wei Chao | Research Scientist | Robot action data and policies | org | C | |
| Zekun Hao | Research Scientist | 3D generative models, VLMs | GANcraft (2021); Meshtron (2024); DualSDF (2020) | org · site | A |
| Zhaoshuo (Max) Li | Research Scientist | 3D understanding, robot policies | Neuralangelo (2023) | org · site | B |
| Zhiqi Li | Research Scientist | Grounding, robot perception | BEVFormer (2022); InternImage (2023); Eagle 2 (2025) | org · site | C |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Natural Language Processing / Foundation Models (Nemotron, Megatron-LM data+modeling) | Mohammad Shoeybi (NLP team manager) with Mostofa Patwary (Director, Large Foundational Language Model) | Megatron-LM (2019); Megatron-Turing NLG 530B (2022) | A | ||
| Computer Vision & Graphics (DLSS, VLM) | Andrew Tao | pix2pixHD (2018); Partial Convolution Inpainting (2018) | A | ||
| Speech & Audio (ADLR-Audio: WaveGlow/RAD-TTS/BigVGAN/Audio Flamingo line) | Rafael Valle led/represented it through ~2025 and has since LEFT NVIDIA; Wei Ping is now the named project lead on the audio-LLM work (UALM), with Zhifeng Kong and Sang-gil Lee as the senior ICs | WaveGlow (2019); BigVGAN (2023) | A | ||
| Alignment / Post-training | Jonathan Raiman | ChatQA (2024); NV-Embed (2024) | A | ||
| Chip Design (AI for EDA: PrefixRL, CircuitVAE, ChipNeMo collaboration, Nemotron-CORTEXA) | Rajarshi Roy / Jialin Song (with Jonathan Raiman) | PrefixRL (2021); ChipNeMo (2023) | A | ||
| Systems for Deep Learning (Megatron-LM training/inference systems) | no single named manager found; Jared Casper, Deepak Narayanan, Vijay Korthikanti and Patrick LeGresley are the bio-confirmed ADLR systems researchers | Megatron-LM (2019); Megatron-LM on GPU Clusters (2021) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Abhinav Khattar | Research Scientist | LLM alignment, MoE | — | C | |
| Adrian Lancucki | Research Scientist (NVIDIA Warsaw) | TTS alignment, model efficiency | RAD-TTS (2021); One TTS Alignment to Rule Them All (2022) | — | C |
| Ali Hatamizadeh | Staff Research Scientist, NVIDIA Research | LLM reasoning, RL pretraining | Gated DeltaNet (2025); RLP (2025) | org · site | C |
| Ambrish Dantrey | Engineer/Research Scientist | Speech denoising | CleanUNet 2 (2023); OmniVinci (2025) | — | C |
| Andrew Tao | Distinguished Engineer; manager, Computer Vision team (ADLR) | Vision, DLSS, VLM | pix2pixHD (2018); Partial Convolution Inpainting (2018); Hierarchical Multi-Scale Attention (2020) | site | A |
| Arushi Goel | Research Scientist | Audio-language models | Audio Flamingo 3 (2025); OMCAT (2024) | — | B |
| Atefeh Sohrabizadeh | Research Scientist | SWE agents, HW-design ML | — | C | |
| Boxin Wang | Research Scientist (ADLR) | Post-training, RL, VLM | DecodingTrust (2023); InstructRetro (2024); NVLM (2024) | site | B |
| Brandon Norick | Research Scientist (ADLR) | Pretraining data, model architecture | Megatron-Turing NLG 530B (2022); An Empirical Study of Mamba-based Language Models (2024) | — | B |
| Bryan Catanzaro | Vice President, Applied Deep Learning Research | Applied DL across NVIDIA; Nemotron | Megatron-LM (2019); cuDNN (2014); Deep Speech 2 (2016) | site | A |
| Chankyu Lee | Research Scientist, ADLR | Retrieval, RAG, post-training | NV-Embed (2024); Nemotron-Cascade (2025) | — | A |
| Chao-Han Huck Yang | Senior Research Scientist, NVIDIA Research | Speech-language modeling, alignment | HyPoradise (2023); Whispering-LLaMA (2023) | org · site | C |
| Dan Su | Research Scientist (ADLR foundation models) | Pretraining data curation | — | B | |
| David Tarjan | Principal Research Scientist (Deep Learning), ADLR | DLSS SR/FG/RR performance lead | SDC-Net (2018) | — | A |
| Deepak Narayanan | Senior Applied Deep Learning Research Scientist, ADLR | LLM training/inference efficiency | PipeDream (2019); Megatron-LM on GPU Clusters (2021); DAWNBench (2017) | site | A |
| Duncan Riach | Deep Learning Engineer | Determinism, model architecture | An Empirical Study of Mamba-based Language Models (2024) | — | C |
| Fitsum Reda | Principal Research Scientist, Deep Imagination Research (DIR) - fo | Generative video/vision | FILM (2022); TryOnDiffusion (2023); SDC-Net (2018) | — | A |
| Guilin Liu | Senior Research Scientist | Vision, partial convolutions | Partial Convolution Inpainting (2018); Video-to-Video Synthesis (2018); Eagle (2025) | site | B |
| Helen Ngo | Research Scientist/Engineer | LLM inference | — | C | |
| Jaehyeon Kim | Research Scientist | TTS, audio LLMs | Audio Flamingo 3 (2025) | site | B |
| Jared Casper | Senior Deep Learning Scientist, ADLR | Megatron-LM training systems | Megatron-LM (2019); Megatron-LM on GPU Clusters (2021) | — | A |
| Jialin Song | Senior Research Scientist, ADLR | ML for hardware design, code LLMs | — | A | |
| Joao Felipe Santos | Research Scientist | Audio restoration, TTS | RAD-MMM (2023); Fugatto (2025) | site | B |
| John Kamalu | Research Scientist (ADLR foundation models) | Pretraining data, FP8 recipe | — | B | |
| Jon Barker | Senior Research Scientist, ADLR | Multimodal, applied DL | NVLM (2024); SDC-Net (2018) | — | A |
| Jonathan Raiman | Senior Research Scientist, ADLR; leads ADLR Alignment team | RL, alignment, AI for systems | NV-Embed (2024); Deep Voice 2 (2017) | — | A |
| Joseph Jennings | Deep Learning Algorithm Engineer | Pretraining data curation | — | B | |
| Jupinder Parmar | Research Scientist (ADLR) | Pretraining data, continued pretraining | site | B | |
| Karan Sapra | Senior Research Scientist, ADLR | Vision, segmentation, VLM data | Hierarchical Multi-Scale Attention (2020); Video Propagation and Label Relaxation (2019) | site | A |
| Keshav Santhanam | Research Scientist/Engineer | LLM inference | Heterogeneity-Aware Cluster Scheduling (2020) | — | C |
| Kevin J. Shih | Research Scientist | TTS alignment, RAD-TTS | RAD-TTS (2021); Flowtron (2020) | — | B |
| Kezhi Kong | Research Scientist, ADLR Foundation Model team | Pretraining data quality/scale | OpenTab (2024); GOAT (2023); VQ-GNN (2021) | site | A |
| Lawrence McAfee | Deep Learning Engineer/Scientist | Megatron-Core, retrieval | Reducing Activation Recomputation (2023); InstructRetro (2024) | — | C |
| Markus Kliegl | Research Scientist (ADLR) | Pretraining data | — | B | |
| Matvei Novikov | Research Scientist | Reasoning data curation | — | C | |
| Mike Chrzanowski | Deep Learning Scientist | FP8 training recipes | Deep Voice (2017) | — | C |
| Mike Ranzinger | Research Scientist | Vision foundation models (RADIO) | RADIOv2.5 (2025) | — | C |
| Mohammad Shoeybi | Senior Director, Applied Research (manages the NLP team within ADL | LLM pretraining, Megatron-LM | Megatron-LM (2019); Megatron-Turing NLG 530B (2022) | site | A |
| Mostofa Patwary | Director, Large Foundational Language Model (ADLR) | LLM pretraining data + scaling | Megatron-LM (2019); Megatron-Turing NLG 530B (2022); Nemotron-4 340B (2024) | site | A |
| Nayeon Lee | Research Scientist (ADLR) | Post-training, VLM alignment | Factuality Enhanced Language Models (2022); NVLM (2024) | — | B |
| Patrick LeGresley | Researcher, NLP group within ADLR | Large-model training systems | Megatron-LM (2019); Megatron-Turing NLG 530B (2022) | — | A |
| Peng Xu | Research Scientist (ADLR) | Long context, RAG | ChatQA 2 (2024); Retrieval Meets Long Context LLMs (2024) | — | C |
| Peter Dykas | Deep Learning Scientist | FP8 / low-precision pretraining | — | C | |
| Rafael Valle | FORMER: Senior Research Scientist and Manager, ADLR-Audio (has sin | Audio/speech generative models | WaveGlow (2019); Flowtron (2020); Fugatto (2025) | site | A |
| Rajarshi Roy | Senior Research Scientist, ADLR | RL for chip design/architecture | NV-Embed (2024); ChatQA (2024) | — | A |
| Robert Kirby | Research Scientist | AI for chip design, speech | SDC-Net (2018) | — | B |
| Roger Waleffe | Applied Deep Learning Research Scientist | Hybrid Mamba-Transformer architecture | An Empirical Study of Mamba-based Language Models (2024) | — | A |
| Rohan Badlani | Research Scientist | Multilingual TTS | One TTS Alignment to Rule Them All (2022); RAD-MMM (2023) | — | B |
| Ryan Prenger | Research Scientist | Neural vocoders (WaveGlow) | WaveGlow (2019); Flowtron (2020) | — | B |
| Sang-gil Lee | Research Scientist, ADLR | Speech/audio generative models | BigVGAN (2023); UALM (2026) | site | A |
| Sanjeev Satheesh | Research Scientist (ADLR foundation models) | Pretraining data, evaluation | — | B | |
| Sheng-Chieh Lin | Research Scientist (ADLR) | Retrieval, post-training | Nemotron-Cascade (2025) | — | C |
| Shrimai Prabhumoye | Senior Research Scientist, ADLR (also Adjunct Prof., Boston Univer | LLM data, reasoning, safety | RLP (2025); Nemotron-CrossThink (2025) | site | A |
| Siddharth Gururani | Research Scientist | Expressive speech synthesis | SPACE (2023); Fugatto (2025); Cosmos 3 (2026) | org · site | C |
| Sungwon Kim | Research Scientist | Zero-shot TTS, flow matching | P-Flow (2023); ETTA (2025) | — | B |
| Syeda Nahida Akter | Research Scientist (ADLR) | Reasoning data, RL pretraining | Nemotron-CrossThink (2025); RLP (2025) | — | B |
| Teodor-Dumitru Ene | Research Scientist | AI for chip design | — | C | |
| Tuomas Rintamaki | Research Scientist | Vision-language models | NVLM (2024) | — | C |
| Vijay Korthikanti | Senior Research Scientist, ADLR | Scalable training, sequence parallelism | Reducing Activation Recomputation (2023); Megatron-LM on GPU Clusters (2021) | — | A |
| Wei Ping | Distinguished Research Scientist & Director | LLM post-training, reasoning, multimodality | Deep Voice 3 (2018); DiffWave (2021); ChatQA (2024) | site | A |
| Wenliang Dai | Research Scientist (ADLR) | Multimodal LLMs | InstructBLIP (2023); NVLM (2024); Nemotron-Cascade (2025) | site | B |
| Yang Chen | Research Scientist (ADLR) | Math reasoning, RL | AceReason-Nemotron (2025); AceMath (2025) | — | B |
| Yejin Choi | Senior Director / Distinguished Scientist, NVIDIA (also Professor, | Reasoning, synthetic data | ATOMIC (2019); COMET (2019); HellaSwag (2019) | site | C |
| Ying Lin | Research Scientist (ADLR foundation models) | Pretraining data, multilingual | site | B | |
| Zhifeng Kong | Senior Research Scientist | Audio foundation models | DiffWave (2021); Audio Flamingo (2024); CleanUNet 2 (2023) | site | A |
| Zhuolin Yang | Research Scientist (ADLR) | Post-training, cascade RL | Nemotron-Cascade (2025); NVLM (2024) | — | B |
| Zihan Liu | Senior Research Scientist, ADLR | Math reasoning, post-training, RAG | ChatQA (2024); AceReason-Nemotron (2025); AceMath (2025) | site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Enze Xie | diffusion, efficient architectures | SegFormer (2021); PVT (2021); SANA (2025) | site | A | |
| Han Cai | efficient architectures, NAS | ProxylessNAS (2019); Once-for-All (2020); EfficientViT (2023) | org · site | A | |
| Ligeng Zhu | efficient training, VILA | Deep Leakage from Gradients (2019); PockEngine (2023); NVILA (2025) | site | A | |
| Song Han | Research Director; leads Efficient AI, co-leads NVIDIA Singapore L | efficient AI computing | Deep Compression (2016); AWQ (2024); SmoothQuant (2023) | org · site | A |
| Yao (Jason) Lu | efficient AI, VLM | VILA (2024); NVILA (2025) | — | A | |
| Yujun Lin | quantization, efficient inference | SVDQuant (2025); QServe (2025); TorchSparse (2022) | site | A | |
| Yukang Chen | long-context, efficient LLM | LongLoRA (2024); LongVILA (2025); VoxelNeXt (2023) | site | A | |
| Yuyang Zhao | 3D and video generation | Animate124 (2024); GenXD (2025); SANA-Video (2026) | site | A | |
| Zhijian Liu | efficient 3D and multimodal | NVILA (2025) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Management line under Pavone | Boris Ivanovic | Alpamayo-1 (2025); trajdata (2023) | A | ||
| Next-Generation AV Architectures (research thrust) | Alpamayo-1 (2025); DiffStack (2022) | E | |||
| AV Foundation Models (research thrust) | Alpamayo-R1 (2025); Wolf dense video captioning (2025) | E | |||
| Simulation (research thrust) | CTG controllable traffic simulation (2023); BITS (2023) | E | |||
| AI Safety for AVs (research thrust) | Sample-Efficient Safety Assurances using Conformal Prediction (2022); Task-Relevant Failure Detection for Trajectory Predictors (2022) | E |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Apoorva Sharma | Research Scientist | uncertainty quantification for safe ML | SCOD: Sketching Curvature for OOD Detection (2021); Continuous Meta-Learning without Tasks (2020) | org · site | A |
| Bertrand Douillard | end-to-end AV models, RSSM, RFT | org | A | ||
| Boris Ivanovic | Senior Research Scientist and Manager | trajectory forecasting, end-to-end AV stacks | Alpamayo-1 (2025); trajdata (2023); DiffStack (2022) | org · site | A |
| Boyi Li | Research Scientist | multimodal, data-efficient vision-language models | Wolf dense video captioning (2025); Describe Anything (2025) | org · site | A |
| Chaowei Xiao | Research Scientist (Faculty Scientist); Assistant Professor, Unive | trustworthy/secure multimodal foundation models | Spatially Transformed Adversarial Examples (2018); DiffPure (2022); AutoDAN (2024) | org · site | A |
| Ed Schmerling | Research Scientist | generative modeling, simulation, safety assurance | Learning Autonomous Vehicle Safety Concepts from Demonstrations (2023); Real-Time Anomaly Detection and Reactive Planning with LLMs (2024) | org · site | A |
| Gerry Che | object-centric learning for driving agents | CTG controllable traffic simulation (2023) | org | A | |
| Heng Yang | Research Scientist (Faculty Scientist); Assistant Professor, Harva | certifiable perception, trustworthy autonomy | Graduated Non-Convexity (GNC) robust estimation (2020); Object Pose Estimation with Statistical Guarantees (2023) | org · site | A |
| Jack (Huize) Shi | Senior Research Engineer | end-to-end autonomous driving | org | A | |
| Jef Packer | Research Engineer | interaction-aware planning and decision-making | org · site | A | |
| Karen Leung | Research Scientist (Faculty Scientist); Assistant Professor, Unive | safe interaction-aware planning, formal methods | Learning Autonomous Vehicle Safety Concepts from Demonstrations (2023); Interaction-Dynamics-Aware Perception Zones (2022) | org · site | A |
| Marco Pavone | Director of Autonomous Vehicle Research, NVIDIA (Associate Profess | autonomous systems design and control | Alpamayo-1 (2025); CTG controllable traffic simulation (2023); BITS (2023) | org · site | A |
| Maximilian Igl | Senior Research Scientist | data-driven AV simulation, RL generalization | org · site | A | |
| Michael Watson | Research Engineer | AV simulation and control algorithms | Sim2Val (2025) | org | A |
| Peter Karkus | Research Scientist | differentiable/structured planning for driving | DiffStack (2022); Differentiable SLAM-net (2021); Differentiable Mapping Networks (2020) | org · site | A |
| Rachel Luo | Research Scientist | uncertainty quantification, AV safety/reliability | Sample-Efficient Safety Assurances using Conformal Prediction (2022); Local Calibration: Metrics and Recalibration (2022) | org · site | A |
| Sushant Veer | Research Scientist | safe decision-making, out-of-ODD edge cases | Receding Horizon Planning with Rule Hierarchies (2023); Task-Relevant Failure Detection for Trajectory Predictors (2022); Learning Autonomous Vehicle Safety Concepts from Demonstrations (2023) | org · site | A |
| Wenhao Ding | Research Scientist | safety-critical scenario generation, causal RL | Alpamayo-R1 (2025); RealDrive (2025); CAT-K closed-loop SFT of tokenized traffic models (2025) | org · site | A |
| Wenjie Luo | Senior Research Scientist | deep perception, prediction and planning | Fast and Furious (2018); PIXOR (2018); End-to-end Interpretable Neural Motion Planner (2019) | org · site | A |
| Xiangyu Chen | Senior Research Engineer | multimodal learning for driving perception | DiffuBox (2024); Waymax (2023) | org · site | A |
| Xinshuo Weng | Research Scientist | 3D multi-object tracking, prediction | MTP: Multi-Hypothesis Tracking and Prediction (2022); Whose Track Is It Anyway? (2022); Robust Trajectory Prediction against Adversarial Attacks (2022) | org · site | A |
| Yan Wang | Research Scientist | cost-effective 3D driving perception | Pseudo-LiDAR (2019); Alpamayo-R1 (2025); STORM (2025) | org · site | A |
| Yulong Cao | Research Scientist | AV perception security and safety | AdvDO (2022); Robust Trajectory Prediction against Adversarial Attacks (2022); You Can't See Me: physical LiDAR removal attacks (2023) | org · site | A |
| Yurong You | Research Scientist | multi-camera encoding, end-to-end driving | Pseudo-LiDAR++ (2020); Hindsight is 20/20 (2022); Ithaca365 (2022) | site | A |
| Yuxiao Chen | Research Scientist | safety-critical planning and decision-making | ScePT (2022); Tree-structured Policy Planning (2023); BITS (2023) | org | A |
| Zheng Lian | Research Engineer | AV data and platform infrastructure | org | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Perception | RVT (2023); RVT-2 (2024) | A | |||
| Task and Motion Planning | cuRobo (2023); cuTAMP / Differentiable GPU-Parallelized TAMP (2024) | A | |||
| Control | cuRobo (2023); Geometric Fabrics (2022) | A | |||
| Reinforcement Learning | SHAC: Accelerated Policy Learning with Parallel Differentiable Simulation (2022); IndustReal (2023) | A | |||
| Imitation Learning | MimicGen (2023); SkillGen (2024) | A | |||
| Simulation | Factory: Fast Contact for Robotic Assembly (2022); DiSECt (2021) | A | |||
| Human-Robot Interaction | Reactive Human-to-Robot Handovers (2021); Inference-Time Policy Steering (2025) | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alexander Lefort | Lab Technician | robot hardware lab support | — | A | |
| Alperen Degirmenci | Senior Robotics Research Engineer | world models, end-to-end learning, VLAs | The NVIDIA PilotNet Experiments (2020) | org · site | A |
| Ankit Goyal | Senior Research Scientist | 3D perception, manipulation policy learning | RVT (2023); SimpleView (2021); Non-Deep Networks / ParNet (2022) | org · site | A |
| Anqi Li | Research Scientist | safe and sample-efficient robot learning | Geometric Fabrics (2022); HAMSTER (2024) | org · site | A |
| Caelan Garrett | Senior Research Scientist | task and motion planning, manipulation | PDDLStream (2020); Integrated Task and Motion Planning (2021); SkillGen (2024) | org · site | A |
| Claudia Pérez D'Arpino | Research Scientist, Robotics-AI | policy steering, VR teleoperation, HRI | iGibson (2021); Inference-Time Policy Steering (2025) | org · site | A |
| Elie Aljalbout | Research Scientist | world models, sim-to-real, RL | The Reality Gap in Robotics (2026) | org · site | A |
| Fabio Ramos | Principal Research Scientist (also Professor, University of Sydney | Bayesian inference, uncertainty, sim-to-real | cuRobo (2023); STORM joint-space MPC (2021); SHAC: Accelerated Policy Learning with Parallel Differentiable Simulation (2022) | org · site | A |
| Hugo Hadfield | Senior Robotics Research Software Engineer | dexterous manipulation, real-time control | org · site | A | |
| Iretiayo Akinola | Senior Research Scientist | accelerated simulators, sim-to-real transfer | TacSL (2024); Factory: Fast Contact for Robotic Assembly (2022); IndustReal (2023) | org · site | A |
| Jie Xu | Research Scientist | differentiable simulation, robot learning | SHAC: Accelerated Policy Learning with Parallel Differentiable Simulation (2022); Neural Robot Dynamics (2025); TacSL (2024) | org · site | A |
| Lars Johannsmeier | Research Scientist | contact-intensive manipulation, deployable systems | Process-centric manipulation taxonomy for tactile robot skills (2025) | org · site | A |
| Moritz Benno Reuss | Research Scientist | efficient robot foundation models, VLAs | FLOWER (2025); MoDE (2025); BESO (2023) | site | A |
| Rowland O'Flaherty | Senior Robotics Research Software Engineer | planning, optimal control, research software | scene_synthesizer (2025) | org | A |
| Tucker Hermans | Faculty Scientist (NVIDIA senior research scientist; Associate Pro | multi-sensory manipulation, learning and control | StructDiffusion (2023); DefGraspSim (2022); DefGraspNets (2023) | org · site | A |
| Wei Yang | Research Scientist | vision for manipulation, human-robot handover | Reactive Human-to-Robot Handovers (2021); HandoverSim (2022); DexYCB (2021) | org · site | A |
| Xuning Yang | Senior Research Scientist | policy evaluation benchmarks, mobile manipulation | RoboLab (2026); RoboArena (2025); Aim My Robot (2025) | org · site | A |
| Yashraj Narang | Team Leader (lab lead) | contact-rich manipulation, sim-to-real | Factory: Fast Contact for Robotic Assembly (2022); IndustReal (2023); DexYCB (2021) | org · site | A |
| Yijie Guo | demonstration-efficient manipulation learning | RVT (2023); RVT-2 (2024) | org · site | A | |
| Yixuan Wang | Research Scientist | robotic manipulation, 3D representations | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Redmond, WA site group | Chris Wyman | ReSTIR (2020); RTXDI (2021) | A | ||
| Helsinki, Finland site | Gradient-Domain Path Tracing (2015); ReSTIR GI (2021) | A | |||
| Lund / Sweden site cluster | Noise2Noise (2018); nvdiffrec (2022) | D | |||
| Zurich / Switzerland site cluster | Neural Importance Sampling (2019); Neural Control Variates (2020) | E | |||
| Research thrusts (from the lab's own homepage text, not org units) | Aaron Lefohn | DLSS; RTXDI | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Aaron Lefohn | Vice President, Graphics Research (RTR Team Leader) | leads real-time graphics research | Real-Time Neural Appearance Models (2023); Recurrent Denoising Autoencoder for Monte Carlo sequences (2017); Aggregate G-Buffer Anti-Aliasing (2015) | org | A |
| Andrea Weidlich | Principal Researcher | appearance and spectral material modeling | Real-Time Neural Appearance Models (2023); Microfacet Theory for Non-Uniform Heightfields (2023) | org | A |
| Bart Wronski | Principal Research Scientist | ML plus signal/image processing | Stochastic Texture Filtering (2024); Random-Access Neural Compression of Material Textures (2023); Collaborative Texture Filtering (2025) | org · site | A |
| Benedikt Bitterli | Senior Research Scientist | neural appearance and scene representation | ReSTIR (2020); Generalized Resampled Importance Sampling (2022); Real-Time Neural Appearance Models (2023) | org · site | A |
| Bing Xu | Research Scientist | physically-based and neural rendering | NeuSample (2023); Residual Path Integrals for Re-Rendering (2024) | org · site | A |
| Blaire Yu | Research Scientist | generative materials, wave-optics appearance | Iridescent Feather Appearance Model (2024); Full-Wave Reference Simulator for Surface Reflectance (2023) | org · site | A |
| Chris Cummings | Senior Rendering Engineer | neural graphics tooling and infrastructure | org | A | |
| Chris Wyman | Research Director (Distinguished Research Scientist), Redmond WA | real-time ray tracing, ReSTIR sampling | ReSTIR (2020); RTXDI (2021); SVGF spatiotemporal variance-guided filtering (2017) | org · site | A |
| Craig Kolb | real-time path tracing systems | Falcor rendering framework (2022) | org | A | |
| Cyril Crassin | Research Scientist | voxel GI, filtering, ray-tracing pipelines | GigaVoxels (2009); Voxel Cone Tracing / VXGI (2011); The SGGX Microflake Distribution (2015) | org · site | A |
| Eugene d'Eon | Principal Research Scientist | volumetric scattering, linear transport theory | Sampling the Distribution of Visible Normals (2014); Efficient Rendering of Human Skin (2007); Quantized-Diffusion Model for Translucent Materials (2011) | org · site | A |
| Fabrice Rousselle | Senior Research Scientist | deep learning for real-time rendering | Neural Importance Sampling (2019); Kernel-Predicting Convolutional Networks denoiser (2017); Neural Radiance Caching (2021) | org · site | A |
| Jacob Munkberg | Principal Research Scientist | differentiable rendering, 3D reconstruction | Noise2Noise (2018); nvdiffrec (2022); FlexiCubes (2023) | org · site | A |
| Jan Novak | Principal Research Scientist | rendering and machine learning | Neural Importance Sampling (2019); Neural Control Variates (2020); Real-Time Neural Radiance Caching (2021) | org · site | A |
| Jon Hasselgren | differentiable rendering and denoising | Noise2Noise (2018); nvdiffrecmc (2022); Conservative Rasterization (2005) | org · site | A | |
| Lifan Wu | Research Scientist | differential light transport, inverse rendering | Appearance-Preserving Prefiltering for Displacement-Mapped Surfaces (2019); Analytic Spherical Harmonic Gradients for Many Polygonal Area Lights (2020); Differentiable Time-Gated Rendering (2021) | org · site | A |
| Markus Kettunen | Senior Research Scientist | Monte Carlo light transport theory | Gradient-Domain Path Tracing (2015); ReSTIR GI (2021); E-LPIPS (2019) | org · site | A |
| Matt Pharr | Distinguished Research Scientist | ray tracing, ML for rendering | Physically Based Rendering / pbrt (2004); ispc (2012); Stochastic Texture Filtering (2024) | org · site | A |
| Milos Hasan | Principal Scientist | neural and procedural material generation | Matrix Row-Column Sampling (2007); MaterialGAN (2020); MatFormer (2022) | org · site | A |
| Peter Kocsis | Research Scientist | generative inverse and forward rendering | Intrinsic Image Diffusion (2024); LightIt (2024); IntrinsiX (2025) | org · site | A |
| Petrik Clarberg | Distinguished Research Scientist | real-time path tracing pipeline research | Wavelet Importance Sampling (2005); Falcor rendering framework (2022); Real-Time Neural Appearance Models (2023) | org · site | A |
| Pontus Ebelin | Research Scientist | image quality metrics, human perception | FLIP image difference evaluator (2020); Temporally Dense Ray Tracing (2019); Collaborative Texture Filtering (2025) | org · site | A |
| Ravi Ramamoorthi | Distinguished Research Scientist, Graphics (part-time); Ronald L. | light transport, sampling, neural rendering | NeRF (2020); Irradiance Environment Maps / spherical harmonic lighting (2001); A Signal-Processing Framework for Inverse Rendering (2001) | org · site | A |
| Saeed Hadadan | Research Scientist | generative rendering, neural materials | Neural Radiosity (2021); Inverse Global Illumination using a Neural Radiometric Prior (2023); Generative Detail Enhancement for Physically Based Materials (2025) | site | A |
| Sai Bangaru | Research Scientist | differentiable compilers, Slang autodiff | Slang.D (2023); Unbiased Warped-Area Sampling for Differentiable Rendering (2020); Teg / Systematically Differentiating Parametric Discontinuities (2021) | org · site | A |
| Simon Kallweit | Senior Software Engineer (title as of joining in 2019) | rendering research software infrastructure | Real-Time Neural Appearance Models (2023); Falcor rendering framework (2022) | org · site | A |
| Steve Marschner | Professor of Computer Science, Cornell University; part-time NVIDI | appearance modeling of materials | Light Scattering from Human Hair Fibers (2003); A Practical Model for Subsurface Light Transport (2001); Manifold Exploration (2012) | org · site | A |
| Tizian Zeltner | Senior Research Scientist | appearance modeling, differentiable rendering | Real-Time Neural Appearance Models (2023); Specular Manifold Sampling (2020); Monte Carlo Estimators for Differential Light Transport (2021) | org · site | A |
| Tomas Akenine-Moller | Distinguished Research Scientist | real-time rendering, ray tracing, hardware | Moller-Trumbore ray-triangle intersection (1997); Real-Time Rendering (book, 1999); FLIP image difference evaluator (2020) | org · site | A |
| Tomas Davidovic | Principal Engineer | real-time rendering, scene processing | Real-Time Neural Appearance Models (2023) | org | A |
| Ziyi Zhang | Research Scientist | light transport, differentiable rendering | Radiance Surfaces (2025); Many-Worlds Inverse Rendering (2025) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alex Trevithick | Research Scientist | generative 3D/4D vision, world models | GRF (2021); Live 3D Portrait (2023); What You See Is What You GAN (2024) | org · site | A |
| Amrita Mazumdar | Senior Research Scientist | streaming, compression, 3D/4D representations | QUEEN (2024); Play4D (2025); AI-Mediated 3D Video Conferencing (2023) | org · site | A |
| Christian Jacobsen | Research Scientist | generative modeling, human-robot interaction | CoCoGen (2025); 3D-Generalist (2025) | org · site | A |
| Jonghyun Kim | Senior Research Scientist | holographic near-eye AR/VR displays | NVGaze (2019); Neural Holography (2020); AI 3D Selfie (2025) | org · site | A |
| Koki Nagano | Principal Research Scientist | digital humans, generative 3D, media forensics | EG3D (2022); Live 3D Portrait (2023); GeNVS (2023) | org · site | A |
| Michael Stengel | Research Scientist | computer graphics, telepresence, VR rendering | NVGaze (2019); Foveated AR (2019); Live 3D Portrait (2023) | org · site | A |
| Seonwook Park | Senior Research Scientist | computational human perception, gaze estimation | Few-Shot Adaptive Gaze Estimation (2019); ETH-XGaze (2020); Towards End-to-end Video-based Eye-tracking (2020) | org · site | A |
| Shalini De Mello | Director of Research, New Experiences | human-AI interaction, 4D world modeling | EG3D (2022); GroupViT (2022); Few-Shot Adaptive Gaze Estimation (2019) | org · site | A |
| Shengze Wang | Research Scientist | lifelike robots, digital humans, telepresence | BLADE (2025); Coherent3D (2025); EDPLVO (2022) | org · site | A |
| Tianye Li | Research Scientist | digital humans, 4D capture, graphics | FLAME (2017); SoftRas (2019); QUEEN (2024) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alex Zook | games, generative agents, 3D worlds | 3D-GENERALIST (2026); FactorSim (2025); Interactive Texture Painting with Generative AI (2024) | org · site | A | |
| Ben Boudaoud | esports, HCI, input latency | Foveated AR (2019); SIDOD (2019); Constant Field of View Display Size Effects on First-Person Aiming Time (2023) | org · site | A | |
| Ekta Prashnani | perceptual quality, deepfake detection | PieAPP (2018); Avatar Fingerprinting (2024); Noise-Aware Video Saliency Prediction (2021) | org · site | A | |
| Jaehyun (Jae-Hyun) Jung | vision science, AR/VR, low vision | Active Confocal Imaging for Visual Prostheses (2015); Multiplexing Prisms for Field Expansion (2017) | org · site | A | |
| Joohwan Kim | Research Manager | applied perception, latency, esports | Towards Foveated Rendering for Gaze-Tracked VR (2016); Foveated AR (2019); NVGaze (2019) | org · site | A |
| Josef Spjut | esports, rendering, player performance | Temporally Dense Ray Tracing (2019); Scaling Probe-Based Real-Time Dynamic Global Illumination for Production (2020); Hybrid Modulation for Near Zero Display Latency (2016) | org · site | A | |
| Peter Xenopoulos | esports analytics, sports ML | Awpy Counter-Strike demo parser; ESTA esports trajectory dataset | org · site | A | |
| Ruth Rosenholtz | peripheral vision, visual perception | Measuring Visual Clutter (2007); Feature Congestion display clutter measure (2005); Capabilities and Limitations of Peripheral Vision (2016) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Assaf Hallak | reinforcement learning | SoftTreeMax (2025); Improve Agents without Retraining: Parallel Tree Search with Off-Policy Correction (2021) | org | A | |
| Chen Tessler | Research Scientist | RL, physics-based character control | Reward Constrained Policy Optimization (2019); CALM (2023); MaskedMimic (2024) | org · site | A |
| Eli Meirom | graph neural networks, RL | From Local Structures to Size Generalization in Graph Neural Networks (2021); Controlling Graph Dynamics with RL and GNNs (2021); 'This is my unicorn, Fluffy' (2022) | org · site | A | |
| Gal Chechik | Sr. Director of AI Research | vision-language, learning theory | Textual Inversion (2022); ConsiStory (2024); Add-it (2025) | org · site | A |
| Gal Dalal | Senior Research Scientist | RL, planning, search | Safe Exploration in Continuous Action Spaces (2018); Finite Sample Analyses for TD(0) with Function Approximation (2018); RL for Datacenter Congestion Control (2022) | org · site | A |
| Haggai Maron | Research Scientist | equivariant deep learning, geometry | On Learning Sets of Symmetric Elements (2020); Equivariant Architectures for Learning in Deep Weight Spaces (2023); Graph Metanetworks (2023) | org · site | A |
| Ido Greenberg | Senior Research Scientist | risk-averse RL, robust control | Efficient Risk-Averse Reinforcement Learning (2022); Optimization or Architecture: How to Hack Kalman Filtering (2023); Train Hard, Fight Easy: Robust Meta RL (2023) | org · site | A |
| Shie Mannor | reinforcement learning theory | Reward Constrained Policy Optimization (2019); CALM (2023); Reinforcement Learning for Datacenter Congestion Control (2022) | org · site | A | |
| Yftah Ziser | NLP, LLM analysis | site | A | ||
| Yoad Tewel | Research Scientist | text-to-image personalization | Perfusion / Key-Locked Rank One Editing (2023); ConsiStory (2024); Add-it (2025) | org · site | A |
| Yoni Kasten | 3D vision, SfM, reconstruction | Layered Neural Atlases (2021); VolSDF: Volume Rendering of Neural Implicit Surfaces (2021); Text2LIVE (2022) | org · site | A | |
| Yuval Atzmon | vision-language, generative models | Textual Inversion (2022); Perfusion / Key-Locked Rank One Editing (2023); ConsiStory (2024) | org | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| ByungKwan Lee | multimodal LLMs | VLsI (2025) | — | A | |
| Cheng Sun | 3D reconstruction, neural fields | DirectVoxGO (2022); HorizonNet (2019); HoHoNet (2021) | org · site | A | |
| Chi-Pin Huang | Research Scientist | video generation, editing | ThinkAct (2025); VideoMage (2025) | org · site | A |
| Fred (Fu-En) Yang | Research Scientist | vision-language, embodied AI | ThinkAct (2025); VideoMage (2025); RAPPER (2024) | org · site | A |
| Guan-Ting (Danny) Liu | LLM agents; also listed under EDA | Hierarchical Programmatic Reinforcement Learning (2023); Sub-Resolution Assist Feature Generation with RL and Transfer Learning (2022) | org · site | A | |
| Jaesung Choe | 3D vision, point clouds | Mosaic3D (2025); Dr. Splat (2025) | org · site | A | |
| Min-Hung Chen | Research Scientist | efficient tuning, multimodal LLM | ThinkAct (2025); Hymba (2025); Omni-RGPT (2025) | org · site | A |
| Ryo Hachiuma | Research Scientist | video understanding, human pose | Structured Keypoint Pooling (2023); Dynamics-Regulated Kinematic Policy for Egocentric Pose Estimation (2021); SANER (2024) | org · site | A |
| Sung-Feng Huang | speech synthesis, voice conversion | Meta-TTS (2022); Audio Word2Vec (2019) | org · site | A | |
| Szu-Wei Fu | speech enhancement, audio ML | MetricGAN (2019); MetricGAN+ (2021); Quality-Net (2018) | org · site | A | |
| Yu-Chiang Frank Wang | Research Director, Deep Learning & Computer Vision | generative AI, multimodal, 3D | ThinkAct (2025); Frido (2023); Paraphrasing Is All You Need for Novel Object Captioning (2023) | org · site | A |
| Yusuke Hirota | vision-language bias, captioning | Quantifying Societal Bias Amplification in Image Captioning (2022); Gender and Racial Bias in VQA Datasets (2022); SANER (2024) | org · site | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Chenhui Deng | Research Scientist | LLMs for RTL, graph learning | Polynormer (2024); ChipAlign (2025); ScaleRTL (2025) | org · site | A |
| Chia-Tung (Mark) Ho | LLM agents for design, analog | VerilogCoder (2025); NVCell 2 (2023); DRC-Coder (2024) | org · site | A | |
| Guan-Ting (Danny) Liu | LLM for chip design | Hierarchical Programmatic Reinforcement Learning (2023); Sub-Resolution Assist Feature Generation with RL and Transfer Learning (2022) | org · site | A | |
| Haoxing (Mark) Ren | Director of Design Automation Research | AI/LLM for chip design | ChipNeMo (2023); VerilogEval (2023); DREAMPlace (2019) | org · site | A |
| Haoyu Yang | Research Scientist | DFM, lithography, layout generation | DeePattern (2019); Generic Lithography Modeling with Dual-band Optics-Inspired Neural Networks (2022); ILILT (2024) | org · site | A |
| Mingjie Liu | Research Scientist | LLMs for RTL / ChipNeMo | ChipNeMo (2023); VerilogEval (2023); CraftRTL (2025) | org · site | D |
| Rongjian Liang | Research Scientist | GPU-accelerated EDA, placement | CircuitOps (2024); ChipNeMo (2023) | org | A |
| Wen-Hao Liu | Principal Research Scientist | routing, global placement | NCTU-GR 2.0 (2013); ISPD 2018 Initial Detailed Routing Contest and Benchmarks (2018); GPU/ML-Enhanced Large Scale Global Routing Contest (2024) | org · site | A |
| Yanqing Zhang | circuit design automation, GNNs | Simba (2019); GRANNITE (2020); MAGNet (2019) | org · site | A | |
| Yi-Chen Lu | physical design, placement, ML for EDA | INSTA (2025); C3PO (2026) | site | A | |
| Yunsheng Bai | graph learning for EDA | SimGNN (2019); GLSearch (2021); AssertionForge (2025) | org · site | A | |
| Zongzhi Yu | ML for EDA | HW-NAS-Bench (2021); EDGE-LLM (2024); Spec2RTL-Agent (2025) | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Gemini (frontier foundation models) | Oriol Vinyals and Noam Shazeer (technical co-leads); Jeff Dean as Chief Scientist; Koray Kavukcuoglu as CTO over the whole model stack. Sergey Brin is credited by name for guidance on 2025 model launches and is reported to be hands-on. | Frontier omni-modal LLM family: pre-training, post-training, multimodality, long context, | Gemini 1.0 (2023); Gemini 1.5 with 1M-token context (2024) | Largest single effort in GDM; Gemini technical reports carry 1,000+ named contributors in the contributions appendix | B |
| Gemini — pre-training | Jack Rae and Sebastian Borgeaud (co-leads) | Data curation, scaling laws, architecture, large-scale TPU training runs | Gopher (2021); Chinchilla (2022) | unknown | D |
| Gemini — post-training / evaluation | Slav Petrov (Senior Director, Gemini co-lead) | Instruction tuning, RLHF/RLAIF, safety tuning, evals and benchmarks | Gemini post-training recipes (2024-2026); Gemini eval suite | unknown | C |
| Gemini — multimodal | Jean-Baptiste Alayrac and colleagues from the Flamingo/PaLI lineage | Native image, video, audio and speech understanding fused into the Gemini backbone | Flamingo (2022); Gemini native multimodality (2023-) | unknown | D |
| Gemini — reasoning / thinking | CONTESTED as of mid-2026. Denny Zhou led Google Brain's reasoning team and is reported to have departed in 2026; Jack Rae has publicly presented the Deep Think line of work. | Chain-of-thought, self-consistency, test-time compute (Deep Think), competition maths and | Chain-of-Thought prompting (2022); Self-Consistency (2023) | unknown | E |
| Science (AI for Science) | Pushmeet Kohli — VP of Research, head of the Science and Strategic Initiatives unit; John Jumper directs the AlphaFold line | Structural biology, genomics, mathematics, algorithm discovery, weather, materials, fusion | AlphaFold 2 (2021); AlphaFold 3 (2024) | unknown (hundreds) | A |
| Isomorphic Labs (spun OUT of DeepMind — NOT a GDM subgroup) | Demis Hassabis (Founder & CEO); Colin Murdoch (President, moved from GDM Chief Business Officer); Max Jaderberg (Chief AI Officer, moved from DeepMind 2023); Sergei Yakneen (CTO) | Drug design on AlphaFold-derived models; partnerships with Eli Lilly and Novartis; first h | AlphaFold 3 co-development (2024); Eli Lilly and Novartis partnerships (2024) | ~150-200 (2025 reporting) | B |
| Robotics / Embodied AI (Gemini Robotics) | Carolina Parada — Senior Director and Head of Robotics | Vision-language-action models, dexterous manipulation, embodied reasoning, on-device robot | RT-1 (2022); RT-2 (2023) | unknown | B |
| Generative Media (Veo, Imagen, Lyria, Nano Banana) | Douglas Eck (Senior Research Director) on the creative/generative-media side; Veo credits are headed by Jeff Donahue, Sander Dieleman, Shlomi Fruchter and Ben Poole among core contributors | Video (Veo), image (Imagen, Nano Banana), music (Lyria), speech; distribution via Flow, Wh | Imagen (2022); MusicLM (2023) | unknown | B |
| World models / open-endedness (Genie, SIMA) | Tim Rocktaschel (Research Director, open-endedness) with Satinder Singh in a senior advisory/leadership role; Jack Parker-Holder and Shlomi Fruchter as project co-leads | Real-time generative interactive environments, embodied agents, open-ended learning | Genie (2024); Genie 2 (2024) | ~30 core contributors named on Genie 3, plus ~70 acknowledged | A |
| AI Safety and Alignment | Anca Dragan — Director / Head of AI Safety and Alignment | Technical alignment, mechanistic interpretability, dangerous-capability evals, scalable ov | Frontier Safety Framework (2024, v2 2025, v3 2025); Gemma Scope sparse autoencoders (2024) | unknown (grew substantially 2024-2026) | A |
| Frontier Safety and Governance / Responsibility / Security | Allan Dafoe (Director of Frontier Safety and Governance); Helen King (Senior Director of Responsibility); Four Flynn (VP of Security) | Frontier Safety Framework, misuse threat modelling, policy, responsible deployment, model- | Frontier Safety Framework v1-v3 (2024-2025); Ethical and social risks of harm from language models (2021) | unknown | C |
| AGI safety / long-horizon research | Shane Legg — co-founder, Chief AGI Scientist | Definitions and levels of AGI, AGI safety agenda, capability forecasting | Universal Intelligence (2007); Levels of AGI (2023) | small | B |
| Core / foundational research (the 'research org proper') | Distributed rather than single-headed: David Silver (RL and search), Raia Hadsell (VP of Research), Satinder Singh (senior research lead), Danilo Rezende (deep generative models), under Koray Kavukcuoglu as CTO and Jeff Dean as Chief Scientist | Reinforcement learning, search and planning, deep generative models, graph/algorithmic rea | AlphaGo (2016); AlphaZero (2018) | unknown | B |
| Google Brain (as a distinct entity) | n/a — dissolved as a named organisation at the 2023-04 merger | NOTHING REMAINS as a distinct entity. Brain's people, its TPU/JAX/Pathways infrastructure | Transformer (2017); TensorFlow (2015) | 0 as a named unit | A |
| Gemini app / product (Google product side, not GDM research) | Josh Woodward — VP, Gemini app and Google Labs (since 2025-04) | Consumer Gemini assistant surface. Sits on the Google product side, but the leadership cha | NotebookLM (2023); Gemini app (2025-) | unknown | B |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Allan Dafoe | Director of Frontier Safety and Governance | AI governance, cooperative AI | AI Governance: A Research Agenda (2018); Open Problems in Cooperative AI (2020); Frontier Safety Framework (2024) | site | A |
| Anca Dragan | Director / Head of AI Safety and Alignment; Associate Professor (o | Alignment, human-AI interaction | Legibility and Predictability of Robot Motion (2013); Cooperative Inverse Reinforcement Learning (2016); Frontier Safety Framework (2024) | site | A |
| Ben Poole | Research Scientist | Diffusion, 3D and video generation | DreamFusion (2022); Score-Based Generative Modeling through SDEs (2021); Veo (2024-2025) | site | A |
| Carolina Parada | Senior Director and Head of Robotics | Vision-language-action robot models | RT-2 (2023); Gemini Robotics (2025); Gemini Robotics 1.5 / ER 1.5 (2025) | org | B |
| Danilo J. Rezende | Senior Director / Director of Research, deep generative models | Deep generative models, physics-inspired ML | Stochastic Backpropagation and Approximate Inference / VAE (2014); Variational Inference with Normalizing Flows (2015); Generative Query Networks (2018) | site | A |
| David Silver | VP of Research / Principal Research Scientist, reinforcement learn | RL, search, self-play | AlphaGo (2016); AlphaZero (2018); MuZero (2020) | site | A |
| Demis Hassabis | Co-founder and CEO, Google DeepMind | AGI strategy, science applications | AlphaGo (2016); AlphaFold 2 (2021); Gemini (2023) | org · site | A |
| Doina Precup | Research Team Lead, Montreal (2026 status UNCERTAIN); Professor, M | Hierarchical RL, options framework | Between MDPs and Semi-MDPs / Options (1999); Eligibility Traces for Off-Policy Policy Evaluation (2000); The Option-Critic Architecture (2017) | site | C |
| Douglas Eck | Senior Research Director | Generative media, music, creativity | Magenta (2016); Music Transformer (2018); MusicLM (2023) | org · site | A |
| Iason Gabriel | Senior Staff Research Scientist, ethics and society | AI ethics, value alignment | Artificial Intelligence, Values and Alignment (2020); The Ethics of Advanced AI Assistants (2024) | site | A |
| Jack Parker-Holder | Senior Research Scientist | Open-ended learning, world models | Human-Timescale Adaptation / AdA (2023); Genie 2 (2024); Genie 3 (2025) | site | A |
| Jack Rae | Principal Research Scientist; pre-training and reasoning lead | Scaling, long context, thinking models | Compressive Transformer (2019); Gopher (2021); Gemini Deep Think (2025) | — | D |
| Jean-Baptiste Alayrac | Staff Research Scientist; Gemini multimodal | Vision-language models, video understanding | Flamingo (2022); Gemini (2023) | site | A |
| Jeff Dean | Chief Scientist, Google DeepMind and Google Research | Systems, scaling, TPU-era architecture | MapReduce (2004); TensorFlow (2015); Pathways (2021) | org · site | A |
| Jeff Donahue | Research Scientist; Veo core contributor / co-lead | Generative video and image models | DeCAF (2014); Adversarial Feature Learning / BiGAN (2017); BigGAN (2019) | site | A |
| John Jumper | Director; AlphaFold lead | Protein structure prediction | AlphaFold 2 (2021); AlphaFold 3 (2024) | site | A |
| Josh Woodward | VP, Gemini app and Google Labs (since 2025-04) | Consumer Gemini assistant | NotebookLM (2023); Gemini app (2025-) | — | B |
| Koray Kavukcuoglu | CTO, Google DeepMind; SVP and Chief AI Architect, Google (since 20 | Model stack, research-to-product integration | Human-level control through deep RL / DQN (2015); WaveNet (2016); Gemini (2023) | — | B |
| Matej Balog | Staff Research Scientist | Program synthesis, algorithm discovery | AlphaTensor (2022); FunSearch (2023); AlphaEvolve (2025) | site | A |
| Max Jaderberg | DEPARTED GDM — Chief AI Officer, Isomorphic Labs (since 2023) | Drug design models | Spatial Transformer Networks (2015); Population Based Training (2017); AlphaStar (2019) | site | A |
| Neel Nanda | Senior Research Scientist; leads the mechanistic interpretability | Mechanistic interpretability | TransformerLens (2022); Progress Measures for Grokking via Mechanistic Interpretability (2023); Gemma Scope (2024) | site | A |
| Noam Shazeer | VP; technical co-lead, Gemini (rejoined Google 2024-08) | Transformer architectures, efficiency, mixture-of-experts | Attention Is All You Need (2017); Sparsely-Gated Mixture-of-Experts (2017); Gemini (2023-) | — | B |
| Oriol Vinyals | VP of Research; technical co-lead, Gemini | Deep learning, sequence models, agents | Sequence to Sequence Learning (2014); Pointer Networks (2015); AlphaStar (2019) | org · site | C |
| Petar Velickovic | Senior Staff Research Scientist (London); Affiliated Lecturer, Uni | Graph neural networks, algorithmic reasoning | Graph Attention Networks (2018); Deep Graph Infomax (2019); Neural Algorithmic Reasoning (2021) | site | A |
| Peter Battaglia | Research Director; physics and weather modelling | Graph networks, simulation, weather | Relational inductive biases, deep learning, and graph networks (2018); GraphCast (2023); WeatherNext (2024) | — | D |
| Pushmeet Kohli | VP of Research; head of the Science and Strategic Initiatives unit | AI for science, algorithm discovery | AlphaFold 2 (2021); AlphaTensor (2022); FunSearch (2023) | org · site | A |
| Quoc V. Le | Distinguished Scientist | Sequence learning, AutoML, scaling | Sequence to Sequence Learning (2014); Neural Architecture Search (2017); EfficientNet (2019) | org · site | A |
| Raia Hadsell | VP of Research | Continual learning, robotics, navigation | Dimensionality Reduction by Learning an Invariant Mapping / DrLIM (2006); Progressive Neural Networks (2016); Elastic Weight Consolidation (2017) | site | A |
| Rohin Shah | Research Scientist; leads an AGI alignment team | Scalable/amplified oversight, alignment agenda | Preferences Implicit in the State of the World (2019); Alignment Newsletter (2018-2022); An Approach to Technical AGI Safety and Security (2025) | site | A |
| Sander Dieleman | Research Scientist / Principal Scientist, generative modelling | Diffusion models, audio and video generation | WaveNet (2016); Veo (2024-2025); 'Diffusion models are autoencoders' blog series | site | A |
| Satinder Singh | Senior research lead / Distinguished Scientist; Professor, Univers | Reinforcement learning, reward design, agents | Reward is Enough (2021); Discovery of Useful Questions as Auxiliary Tasks (2019); Genie 3 guidance (2025) | — | B |
| Sebastian Borgeaud | Research Scientist; Gemini pre-training co-lead | Retrieval, scaling laws, pre-training | RETRO (2021); Chinchilla (2022); Gemini (2023) | — | D |
| Shane Legg | Co-founder and Chief AGI Scientist | AGI definition, timelines, safety | Universal Intelligence (2007); Machine Super Intelligence (2008); Levels of AGI (2023) | site | B |
| Shlomi Fruchter | Research lead, Genie / generative video | World models, video generation | Genie 3 (2025); Veo (2024-2025) | — | A |
| Slav Petrov | Senior Director; Gemini co-lead (post-training and evaluation) | NLP, evaluation, post-training | Learning Accurate, Compact and Interpretable Tree Annotation (2006); Universal Dependencies (2016); Gemini (2023-) | org · site | C |
| Tim Brooks | Research Scientist, video generation and world simulators (joined | Video generation, world simulation | InstructPix2Pix (2023); Sora (2024) | site | A |
| Tim Rocktaschel | Research Director / Senior Staff Research Scientist, open-endednes | Open-endedness, world models, agents | NetHack Learning Environment (2020); Genie (2024); Genie 3 (2025) | site | A |
| Tim Salimans | Principal Research Scientist (Amsterdam) | Diffusion models, distillation, fast sampling | Improved Techniques for Training GANs (2016); PixelCNN++ (2017); Progressive Distillation for Fast Sampling of Diffusion Models (2022) | — | D |
| Volodymyr Mnih | Research Scientist (Toronto) | Deep reinforcement learning | Playing Atari with Deep RL / DQN (2013); Human-level control through deep RL (2015); A3C (2016) | — | D |
| Yee Whye Teh | Senior Research Scientist (part-time); Professor of Statistical Ma | Bayesian ML, probabilistic deep learning | Hierarchical Dirichlet Processes (2006); A Fast Learning Algorithm for Deep Belief Nets (2006); Neural Processes (2018) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Algorithms & Optimization | Vahab S. Mirrokni | Graph mining, large-scale optimization, market algorithms, learning theory; the algorithmi | Online Stochastic Matching: Beating 1-1/e (2009); DeepWalk (2014, Perozzi) | several dozen research scientists across NY/Zurich/Mountain View (est.) | A |
| Graph Mining | Vahab Mirrokni (org lead); Bryan Perozzi, Silvio Lattanzi, Alessandro Epasto, Jonathan Halcrow among the senior scientists | Graph learning, clustering, GNNs at web scale, TensorFlow-GNN | DeepWalk (2014); TensorFlow-GNN (2021) | ~15-25 named researchers on the team page | A |
| Market Algorithms / Large-Scale Optimization / Operations Research | Vahab Mirrokni (umbrella); Sreenivas Gollapudi, Renato Paes Leme, Santiago Balseiro, David Applegate, Laurent Perron (OR-Tools) named on the pages | Auction and mechanism design, LP/MIP solvers at Google scale, constraint programming | OR-Tools (open source, ongoing); PDLP first-order LP solver (2021) | unknown | A |
| System Performance / Systems Research | Parthasarathy (Partha) Ranganathan | Warehouse-scale computing, silicon/TPU co-design, ML for systems, compilers, ML for code | The Datacenter as a Computer (book, 3rd ed. 2018); Google's Video Coding Unit / VCU ASIC (2021) | unknown; ~20 named on the team page | A |
| Climate & Sustainability / Crisis Resilience | Yossi Matias is personally listed on the team page; John Platt is the Google Fellow technical lead for Climate and Science; Grey Nearing leads flood/hydrology | Flood forecasting, wildfire detection, contrail avoidance, weather/climate ML, geothermal | Flood Hub / global AI flood forecasting (Nature 2024); FireSat & wildfire boundary detection (2023-2025) | unknown; ~20 named on the team page | A |
| Applied Science | John Platt (Google Fellow, technical leader for Climate and Science) | AI for the natural sciences — connectomics, genomics, materials, weather, neuroscience | Flood-Filling Networks / H01 human cortex connectome (2021, Jain & Januszewski); DeepVariant (2018, McLean et al.) | unknown; ~25 named on the team page | A |
| Health AI | Greg Corrado is the long-standing Google 'Head of Health AI'; the foundation-model side (AMIE, Med-PaLM, MedGemma) is now led by Alan Karthikesalingam and Vivek Natarajan, who both describe themselves as Google DeepMind | Medical LLMs, diagnostic dialogue, dermatology/retina imaging, wearables & sensor models, | Med-PaLM / Med-PaLM 2 (2022-2023); AMIE diagnostic dialogue agent (Nature 2025) | unknown | B |
| Google Research India (Bangalore) | No single public lab head confirmed for 2026; Manish Gupta was the founding Director (2019-). Team leads: Prateek Jain (ML & Optimization), Partha Talukdar (NLU), Pradeep Shenoy (CogML), Varun Gulshan (Earth Observation Sciences), Aravindan Raghuveer | Indian-language NLP, on-device/efficient ML, ML for societal impact (public health, agricu | MuRIL multilingual Indic model (2021); MatFormer / nested Transformers (2023) | ~100+ (est., unverified) | B |
| Quantum AI | Hartmut Neven | Superconducting qubits, quantum error correction, quantum algorithms & applications | Sycamore quantum supremacy (Nature 2019); Willow chip / below-threshold error correction (Nature 2024) | ~200-300 (est., unverified) | A |
| Responsible AI (residual) | no single public lead in Google Research post-2024 | Fairness, evaluation, data-centric AI, sociotechnical harms — the model-facing part moved | Model Cards (2019); Data Cards / Data Cards Playbook (2022, Pushkarna) | ~30 named on the team page (mix of Research and GDM affiliations) | B |
| Athena | no named lead on the page; Sanjiv Kumar, Don Metzler, Ed H. Chi and Mehryar Mohri are the most senior names listed | Self-described as 'an international team of research scientists and engineers who tackle p | UL2 (2022); ScaNN vector search (2020) | ~40 named on the page | B |
| Security, Privacy and Abuse Prevention | not publicly named | Differential privacy, federated analytics, anti-abuse, ML security | Deep Learning with Differential Privacy (2016, Abadi et al.); Google's differential privacy library (open source, 2019) | unknown | A |
| Networking (Global networking, Cloud networking, Network infrastructure) | Amin Vahdat is the senior Google figure over AI infrastructure and the historical networking VP | Datacenter and WAN networking, optical circuit switching, AI-infrastructure fabric | B4 software-defined WAN (2013); Jupiter Rising datacenter fabric (2015) | unknown | A |
| Paradigms of Intelligence (Pi) | Blaise Agüera y Arcas | Neural computing foundations, active inference, sociality, evolution, Artificial Life — an | Federated Learning (2016-17); Computational Life: self-replicating programs (2024) | small (tens) | A |
| Software Engineering & Programming Languages / ML for code | not publicly named; Danny Tarlow is the most senior ML-for-code scientist | Learned developer tooling, program repair, build systems, ML for compilers | Gated Graph Sequence Neural Networks (2016); DIDACT / learning from the software-engineering process (2023) | unknown | B |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alessandro Epasto | Research Scientist, Google Research NY (Graph Mining) | Graph mining, privacy, ego-networks | Ego-splitting framework (2017); Differentially private clustering (2021) | org | A |
| Amin Vahdat | Google Fellow and Chief Technologist for AI Infrastructure | TPUs, datacenter networks, AI infrastructure | B4 software-defined WAN (2013); Jupiter Rising (2015); TPU v4 with optical circuit switching (2023) | org · site | A |
| Aravindan Raghuveer | Research lead, Google Research India | ML for societal impact, data systems | org | C | |
| Avinatan Hassidim | VP of Engineering, Google Research (Israel) | Algorithms, LLM efficiency, Israel research site | Fast Inference from Transformers via Speculative Decoding (ICML 2023); Flood forecasting (2022-2024) | org | C |
| Blaise Agüera y Arcas | VP and Fellow, Google; CTO of Technology & Society; founder of Par | Foundations of intelligence, artificial life | Federated Learning (2016-17); Computational Life: self-replicating programs (2024); What Is Intelligence? (book, 2025) | org | A |
| Bryan Perozzi | Research Scientist, Google Research (Graph Mining) | Graph representation learning, GNNs | DeepWalk (2014); TensorFlow-GNN (2021); Grale (2020) | org | A |
| Cory McLean | Staff Software Engineer / Research Scientist, Google Research (Gen | Genomics, variant calling | DeepVariant (Nature Biotech 2018); DeepConsensus (2022) | org | A |
| Danny Tarlow | Research Scientist, Google Research Montreal (ML for code) | Machine learning for software engineering | Gated Graph Sequence Neural Networks (2016); Learning to Represent Programs with Graphs (2018); DIDACT (2023) | org · site | A |
| Don Metzler | Senior Staff Research Scientist, Google Research (Athena) | Information retrieval, LLMs for search | Rethinking Search: Model-based Information Retrieval (2021); UL2 (2022); Long Range Arena (2021) | org | B |
| Edith Cohen | Research Scientist, Google Research; Professor at Tel Aviv Univers | Sketching, sampling, differential privacy | Size-estimation framework / all-distances sketches (1997); Composable sketches (2019) | org · site | A |
| Emily Denton | Senior Research Scientist, Google Research — Responsible AI / Tech | Sociotechnical AI, generative media ethics | Deep Generative Image Models using a Laplacian Pyramid (LAPGAN, 2015); Characterising Bias in Vision Datasets (2021) | org | A |
| Greg Corrado | Distinguished Scientist; Head of Health AI, Google (co-founder of | Health AI, neuroscience-inspired ML | DistBelief (2012); word2vec (2013); Diabetic retinopathy screening (JAMA 2016) | org | B |
| Grey Nearing | Research Scientist, Google Research — flood forecasting / hydrolog | AI hydrology, global flood forecasting | Global prediction of extreme floods in ungauged watersheds (Nature 2024); Flood Hub (2023) | org | A |
| James Manyika | SVP of Research, Labs, Technology & Society, Google & Alphabet | Executive owner of Google Research | McKinsey Global Institute AI & automation reports (2017-2018); AI2050 fellowship programme (2022-) | org · site | A |
| Jeff Dean | Chief Scientist, Google DeepMind and Google Research | Chief scientist both orgs | MapReduce (2004); TensorFlow (2016); Pathways / PaLM (2022) | org · site | B |
| John C. Platt | Google Fellow; technical leader for Climate and Science | AI for climate and physical science | SMO algorithm for SVMs (1998); Platt scaling / probabilistic outputs (1999); Contrail avoidance & flood forecasting work (2022-2024) | org | A |
| Karthikeyan Shanmugam | Research Scientist, Google Research India | Causal inference, bandits, learning theory | Learning causal graphs with small interventions (2015); Causal bandits work (2019-2023) | org | A |
| Katherine Heller | Research Scientist, Google Research — health & responsible AI | Bayesian ML, health equity | Bayesian Hierarchical Clustering (2005); Health equity / underspecification work (2021-2023) | org | B |
| Lora Aroyo | Research Scientist, Google Research — data-centric / responsible A | Data quality, rater disagreement, evaluation | Truth is a Lie: Crowd Truth (2015); DICES dataset (2023); Adversarial Nibbler (2023) | org · site | A |
| Mahima Pushkarna | Senior Interaction Designer / Researcher, Google Research — Respon | Data documentation and transparency | Data Cards (2022); Data Cards Playbook (2022) | org | B |
| Martín Abadi | Distinguished Scientist / Research Scientist, Google | Systems security, differential privacy, TensorFlow | TensorFlow (OSDI 2016); Deep Learning with Differential Privacy (CCS 2016); A Calculus for Access Control (1993) | org | C |
| Mehryar Mohri | Research Director / Senior Research Scientist, Google Research NY; | Learning theory, ensembles, ASR algorithms | Foundations of Machine Learning (book, 2012); AdaNet (2017); OpenFst (2007) | org · site | B |
| Michal Januszewski | Research Scientist, Google Research (Connectomics) | Connectomics reconstruction algorithms | Flood-Filling Networks (2018); H01 human cortex connectome (2024) | org | B |
| Milad Hashemi | Research Scientist, Google Research (System Performance) | ML for computer architecture | Learning Memory Access Patterns (2018); A Hierarchical Neural Model of Data Prefetching (2021) | org | A |
| Mingxing Tan | Research Scientist / Staff Software Engineer, Google (perception & | Efficient neural architectures | EfficientNet (2019); EfficientDet (2020); MnasNet (2019) | org | B |
| Partha Talukdar | Senior Staff Research Scientist, Google Research Bangalore — leads | Multilingual NLU, knowledge graphs | MuRIL multilingual Indic model (2021); NELL / never-ending language learning (2010-2015); IndicGenBench (2024) | org · site | A |
| Parthasarathy (Partha) Ranganathan | VP and Technical Fellow, Google (systems and silicon) | Warehouse-scale computing, custom silicon | The Datacenter as a Computer (book, 3rd ed. 2018); Google Video Coding Unit / VCU (2021); Warehouse-scale profiling & 'datacenter tax' (2015) | org | B |
| Pradeep Shenoy | Research Scientist, Google Research India — leads Cognitive Modeli | Cognitive modelling, ML robustness | org | B | |
| Praneeth Netrapalli | Research Scientist, Google Research India | Optimization theory, robust ML | Non-convex matrix completion by alternating minimization (2013); How to escape saddle points efficiently (2017) | org | A |
| Prateek Jain | Research Scientist, Google Research India (ML & Optimization group | Non-convex optimization, efficient on-device ML | Low-rank matrix completion via alternating minimization (2013); FastGRNN (2018); MatFormer (2023) | org · site | A |
| Rahul Sukthankar | Distinguished Scientist, Google Research (machine perception / vid | Video understanding, computer vision | Large-scale Video Classification with CNNs (2014); AVA action dataset (2018); Sports-1M (2014) | org · site | B |
| Sanjiv Kumar | VP, Google Research (New York) — large-scale ML and learning theor | Large-scale ML, retrieval, federated optimization | ScaNN / anisotropic vector quantization (2020); Adaptive Federated Optimization (2021); Are Transformers universal approximators? (2020) | org · site | C |
| Sergei Vassilvitskii | Research Scientist, Google Research (Algorithms) | Approximation algorithms, clustering, algorithms with predictions | k-means++ (2007); Scalable k-means++ (2012); Algorithms with Predictions (book chapter, 2020) | org · site | A |
| Silvio Lattanzi | Research Scientist, Google Research Zurich (Graph Mining) | Graph clustering, sublinear algorithms | Affinity Clustering (2017); Distributed Balanced Partitioning (2016); Consistent k-clustering (2019) | org | A |
| Vahab S. Mirrokni | Google Fellow and VP, Google Research (Algorithms & Optimization) | Graph mining, market algorithms, optimization | Online Stochastic Matching: Beating 1-1/e (2009); Distributed Balanced Partitioning via Linear Embedding (2016); Tight approximation for submodular maximization (2008) | org | A |
| Varun Gulshan | Research Scientist, Google Research India — leads the Earth Observ | Geospatial ML for climate mitigation | Diabetic retinopathy deep learning (JAMA 2016); Earth observation / land-cover models (2022-) | org | A |
| Vinodkumar Prabhakaran | Senior Research Scientist, Google Research — Responsible AI | Fairness, cultural bias in LLMs | Cultural bias / SeeGULL stereotype benchmark (2023); Re-imagining Algorithmic Fairness in India (2021) | org | A |
| Viren Jain | Senior Staff Research Scientist, Google Research — leads the conne | Connectomics, large-scale 3D image segmentation | Flood-Filling Networks (Nature Methods 2018); H01 human cortex connectome (Science 2024); FlyWire / full fly brain connectome (Nature 2024) | org · site | A |
| Yossi Matias | Vice President, Google; Head / General Manager of Google Research | Leads all of Google Research | Google Duplex (2018); Flood Hub / AI flood forecasting (2018-); Speculative decoding (2023) | org | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Quantum Hardware (Santa Barbara) | Michel Devoret (Chief Scientist for Quantum Hardware); Julian Kelly (hardware director) | Superconducting qubit fabrication, error correction, Willow-class processors | Sycamore (2019); Willow / below-threshold surface code (2024) | unknown | C |
| Quantum Algorithms & Applications / Theory | Sergio Boixo (chief scientist, quantum computing theory) — not re-verified for 2026 | Quantum algorithms, quantum chemistry, verifiable advantage benchmarks | Quantum supremacy benchmark design (2018-2019); Quantum Echoes / OTOC advantage (2025) | unknown | C |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Hartmut Neven | VP of Engineering, Google; founder and manager, Quantum AI lab | Quantum computing for machine intelligence | Quantum supremacy using a programmable superconducting processor (2019); Willow chip (2024); Quantum Echoes (2025) | org | A |
| Julian Kelly | Director of Quantum Hardware, Google Quantum AI | Qubit hardware and error correction | Suppressing quantum errors by scaling a surface code logical qubit (Nature 2023); Willow (2024) | org | C |
| Michel Devoret | Chief Scientist for Quantum Hardware, Google Quantum AI; Professor | Superconducting qubit hardware | Transmon qubit (2007); Fluxonium qubit (2009); Quantronium (2002) | org | B |
| Sergio Boixo | Chief Scientist, Quantum Computing Theory, Google Quantum AI | Quantum advantage benchmarks and algorithms | Characterizing quantum supremacy in near-term devices (Nature Physics 2018); Quantum supremacy (Nature 2019) | org | C |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Cloud AI Research | Tomas Pfister | Self-described: 'groundbreaking research with the goal to infuse AI into Google Cloud prod | TabNet (AAAI 2021); Distilling step-by-step (2023) | ~20 named researchers on the team page | A |
| Vertex AI platform & Gemini Enterprise (product engineering) | Saurabh Tiwary (VP/GM); Raj Pai (VP, Product Management, Cloud AI) | Model Garden, Agent Builder / ADK, tuning & serving, Gemini Enterprise, Kaggle, Colab | Vertex AI (2021); Agent Development Kit / A2A protocol (2025) | unknown (product org, hundreds+) | C |
| AI Infrastructure (TPU / systems, shared with Google Research) | Amin Vahdat | Custom silicon, datacenter, network and supply chain for AI serving and training | TPU v4 with optical circuit switches (ISCA 2023); Jupiter Rising (2015) | unknown | A |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Amin Vahdat | Google Fellow and Chief Technologist for AI Infrastructure | TPUs, datacenter networking for AI | Jupiter Rising (2015); TPU v4 optical circuit switching (2023) | org · site | A |
| Chen-Yu Lee | Staff Research Scientist, Cloud AI Research | Agents, distillation, document AI | Deeply-Supervised Nets (AISTATS 2015); Distilling step-by-step (2023); Chain-of-Table (2024) | org | B |
| Hamid Palangi | Research Scientist, Cloud AI Research (ex-Microsoft Research) | Multimodal reasoning, agent evaluation | Deep Sentence Embedding Using LSTM (2016); Multimodal reasoning benchmarks (2023-2025) | org | B |
| Jinsung Yoon | Research Scientist, Cloud AI Research | Tabular/time-series ML, synthetic data | GAIN: Missing Data Imputation using GANs (ICML 2018); TimeGAN (2019) | org | B |
| Raj Pai | Vice President, Product Management, Cloud AI (ex-AWS, led Amazon E | Cloud AI product management | — | C | |
| Saurabh Tiwary | VP & General Manager, Cloud AI, Google | Vertex AI, Gemini Enterprise, Kaggle, Colab | — | C | |
| Sercan O. Arik | Research Scientist / Manager, Cloud AI Research | Tabular deep learning, time series, agents | TabNet (2021); Temporal Fusion Transformers (2021); Deep Voice (2017) | org | A |
| Thomas Kurian | CEO, Google Cloud | Runs Google Cloud | org | A | |
| Tomas Pfister | Head of Cloud AI Research, Google Cloud | Agentic AI, tabular ML, distillation | TabNet (2021); Distilling step-by-step (2023); MLE-STAR (2025) | org · site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Perception / self-supervised vision (Paris + Menlo Park) | Piotr Bojanowski / Patrick Labatut (senior ICs; no confirmed named manager) | Self-supervised image representation learning, foundation vision backbones, dense features | DINO (2021); DINOv2 (2023) | ~25-30 (from DINOv3 author list) | D |
| World models / JEPA (Montreal + Menlo Park + NYC) | Nicolas Ballas / Michael Rabbat (senior; previously under LeCun) | Joint-embedding predictive architectures, video world models, planning and robotics from v | I-JEPA (2023); V-JEPA (2024) | ~30 (from V-JEPA 2 author list) | D |
| FAIR CodeGen | Gabriel Synnaeve | Code generation, code world models, neural compilers, RL for programming | Code Llama (2023); CWM / Code World Model (2025) | ~50 named contributors on CWM | D |
| Audio / speech | Jade Copet / Yossi Adi (senior ICs) | Neural audio codecs, music and speech generation, speech representation learning | wav2vec 2.0 (2020); EnCodec (2022) | Unknown | D |
| Segmentation / video perception | Christoph Feichtenhofer (senior IC) | Promptable segmentation, video understanding | SlowFast (2019); SAM (2023) | Unknown | D |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Adrien Bardes | Research Scientist, Paris | Non-contrastive SSL | VICReg (2022); V-JEPA (2024); V-JEPA 2 (2025) | — | D |
| Andrea Vedaldi | Research Scientist (joint with University of Oxford) | 3D vision, visual geometry | VLFeat (2008); Co-tracker (2023); DINOv3 (2025) | site | D |
| Camille Couprie | Research Scientist, Paris | Vision, geospatial ML | DINOv3 (2025) | — | D |
| Chris Cummins | Research Scientist | ML for compilers | CompilerGym (2022); LLM Compiler (2024); CWM (2025) | — | D |
| Christoph Feichtenhofer | Research Scientist Manager | Video understanding, segmentation | SlowFast (2019); SAM 2 (2024) | site | C |
| Daniel Haziza | Research Engineer | Efficient attention kernels | xFormers (2022); DINOv3 (2025); CWM (2025) | — | D |
| Fabian Gloeckle | Research Scientist, Paris | Code LLMs, multi-token prediction | Multi-token prediction (2024); CWM (2025); DecompRL (2026) | — | D |
| Federico Baldassarre | Research Scientist, Paris | Self-supervised vision | DINOv3 (2025) | — | D |
| Francisco Massa | Research Engineer | Vision systems, kernels | DETR (2020); xFormers (2022); DINOv3 (2025) | — | D |
| Francois Fleuret | Research Scientist (joint with University of Geneva) | Deep learning, efficiency | The Little Book of Deep Learning (2023); CWM (2025) | site | D |
| Franziska Meier | Research Scientist Manager | Robot learning | V-JEPA 2 (2025) | — | D |
| Gabriel Synnaeve | Research Director, FAIR CodeGen, Paris | Code generation, RL | Code Llama (2023); CWM / Code World Model (2025); DecompRL (2026) | site | D |
| Herve Jegou | Research Director, Paris | Similarity search, vision | FAISS (2017); DeiT (2021); DINOv3 (2025) | — | D |
| Hugh Leather | Research Scientist | ML for compilers | CompilerGym (2022); LLM Compiler (2024); CWM (2025) | — | D |
| Huy V. Vo | Research Scientist, Paris | Object discovery, data curation | DINOv3 (2025) | — | D |
| Jacob Kahn | Research Engineer | Speech/ML systems | Flashlight / wav2letter++ (2018); CWM (2025) | — | D |
| Jade Copet | Research Manager, Paris | Audio generation, code | MusicGen / AudioCraft (2023); CWM (2025) | — | D |
| Joelle Pineau | FORMER VP of AI Research, head of FAIR 2023-2025 - last day 2025-0 | RL, health AI, open science | Reproducibility checklist (2018); FAIR open-science / ParlAI programme | site | A |
| Jonas Gehring | Research Scientist, Paris | Sequence models, RL for code | ConvS2S (2017); CWM (2025) | — | D |
| Julien Mairal | Research Scientist (joint with Inria Grenoble) | Optimization, SSL vision | SPAMS sparse coding toolbox (2010); DINOv3 (2025) | site | D |
| Koustuv Sinha | Research Scientist, Montreal | Reasoning, reproducibility, video | V-JEPA 2 (2025) | — | D |
| Marc Szafraniec | Research Engineer | SSL vision, video | DINOv3 (2025); V-JEPA 2 (2025) | — | D |
| Maxime Oquab | Research Scientist, Paris | Self-supervised vision | DINOv2 (2023); DINOv3 (2025) | — | D |
| Maximilian Seitzer | Research Scientist | Object-centric / SSL vision | DINOv3 (2025) | — | D |
| Michael Rabbat | Research Scientist Manager, Montreal | Distributed optimization, SSL | V-JEPA 2 (2025); I-JEPA (2023) | — | D |
| Mido Assran | Research Scientist Manager, Montreal | World models, SSL | I-JEPA (2023); V-JEPA 2 (2025) | site | D |
| Naila Murray | Research Manager | Computer vision | CWM (2025) | — | D |
| Nicolas Ballas | Research Scientist, Montreal | Self-supervised video models | I-JEPA (2023); V-JEPA (2024); V-JEPA 2 (2025) | — | D |
| Oriane Simeoni | Research Scientist, Paris | Unsupervised object discovery | LOST (2021); DINOv3 (2025) | — | D |
| Patrick Labatut | Research Engineering Manager, Paris | SSL vision engineering | DINOv2 (2023); DINOv3 (2025); V-JEPA 2 (2025) | — | D |
| Peter O'Hearn | Research Scientist (joint with UCL) | Program analysis, verification | Separation logic (2001); Infer static analyser (2015); CWM (2025) | — | D |
| Piotr Bojanowski | Research Scientist / Manager, Paris | Self-supervised vision | fastText (2017); DINOv2 (2023); DINOv3 (2025) | — | D |
| Quentin Garrido | Research Scientist, Paris | SSL theory, video | V-JEPA 2 (2025) | — | D |
| Rob Fergus | VP of AI Research; head of FAIR (appointed 2025-05-08) | Research direction, deep learning | ZFNet (2013); Intriguing properties of neural networks (2014) | site | A |
| Taco Cohen | Research Scientist | Equivariance, code world models | Group Equivariant CNNs (2016); CWM (2025); DecompRL (2026) | site | D |
| Timothee Darcet | Research Scientist / PhD, Paris | Vision transformer representations | Vision Transformers Need Registers (2023); DINOv2 (2023); DINOv3 (2025) | — | D |
| Vasil Khalidov | Research Engineer | SSL vision, video | DINOv3 (2025); V-JEPA 2 (2025) | — | D |
| Yann LeCun | FORMER VP & Chief AI Scientist, FAIR founder - departed 2025-11 | World models, SSL | Convolutional networks / LeNet (1989-1998); I-JEPA (2023); V-JEPA 2 (2025) | site | A |
| Yossi Adi | Research Scientist (joint with Hebrew University) | Speech and audio generation | EnCodec (2022); AudioCraft (2023); CWM (2025) | — | D |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| TBD Lab | Alexandr Wang | Frontier/foundation model training - the elite team that trains Meta's flagship LLMs. The | Muse Spark (2026-04); Muse Spark 1.1 (2026-07) | 'A few dozen researchers and engineers' - Meta CFO Susan Li, Sept 2025. Explicitly exempted from the Oct 2025 cuts and still hiring while the rest of MSL was frozen. | B |
| FAIR (Fundamental AI Research) | Rob Fergus | Long-horizon open research: self-supervised visual representations, world models, speech/a | DINOv3 (2025); V-JEPA 2 (2025) | Several hundred; materially reduced by the Oct 2025 cuts | B |
| Products & Applied Research | Nat Friedman | Consumer integration - Meta AI assistant, AI in Facebook/Instagram/WhatsApp, Ray-Ban glass | Meta AI assistant (2024-2026); Meta Model API (2026) | Largest of the four groups (unquantified) | B |
| MSL Infra | Aparna Ramani | Training and serving infrastructure, compilers, GPU fleet. NOTE: as of Jan 2026 the larges | Meta AI training infrastructure / Prometheus and Hyperion clusters (2025-2026) | Unknown | B |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Ahmad Al-Dahle | VP, GenAI (pre-MSL head of the Llama/GenAI organisation) | Generative AI products/models | Llama 3 (2024); Llama 4 (2025) | — | E |
| Alexander Kolesnikov | Researcher, TBD Lab (joined 2025 from OpenAI Zurich; previously Go | Vision-language pretraining | Vision Transformer / ViT (2020); Big Transfer / BiT (2020) | — | C |
| Alexandr Wang | Chief AI Officer, Meta; head of Meta Superintelligence Labs; lead | Frontier model strategy, org lead | Scale AI (2016, co-founder/CEO) | — | A |
| Andrew Bosworth | CTO, Meta; head of Reality Labs (NOT an MSL role) | Hardware, AR/VR, AI devices | Ray-Ban Meta smart glasses (2023-2025); Meta Ray-Ban Display (2025) | — | B |
| Andrew Tulloch | Engineer/researcher (joined Meta 2025-10 from Thinking Machines La | Training systems, model efficiency | PyTorch mobile/quantization work (2018-2021) | — | C |
| Anton Bakhtin | Researcher, TBD Lab (joined 2025 from Anthropic; previously FAIR) | RL, planning, LLM training | CICERO (2022); Diplomacy no-press agent (2021) | — | C |
| Aparna Ramani | VP of Engineering; head of MSL Infra | AI infrastructure, data platforms | — | B | |
| Avi Verma | FORMER researcher - joined mid-2025 from OpenAI, returned to OpenA | LLM research | — | C | |
| Chaya Nayak | FORMER Director of Product Management, GenAI - left for OpenAI, Au | GenAI product | — | C | |
| Daniel Gross | VP of Product, Meta; co-lead of Meta Compute (from 2026-01). Joine | Compute capacity strategy, supplier partnerships | Safe Superintelligence Inc. (2024, co-founder); NFDG venture fund (2022) | — | B |
| Ethan Knight | FORMER researcher - joined mid-2025 from OpenAI, returned to OpenA | LLM research | — | C | |
| Hongyu Ren | Researcher, TBD Lab (joined 2025 from OpenAI) | Post-training, reasoning models | OpenAI o1-mini (2024, contributor) | — | C |
| Huiwen Chang | Researcher, TBD Lab (joined 2025 from OpenAI) | Image generation | MaskGIT (2022); Muse text-to-image (2023) | — | C |
| Jack Rae | Researcher, TBD Lab (joined 2025 from Google DeepMind) | Reasoning, pretraining | Gopher (2021); Chinchilla (2022, co-author) | — | C |
| Ji Lin | Researcher, TBD Lab (joined 2025 from OpenAI) | Efficient inference, multimodal | AWQ (2023); Temporal Shift Module / TSM (2019) | — | C |
| Jiahui Yu | Researcher, TBD Lab (joined 2025 from OpenAI) | Multimodal perception | CoCa (2022); Parti (2022) | — | C |
| Joel Pobar | Engineer, MSL Infra (joined 2025 from Anthropic) | Inference/runtime infrastructure | HHVM (2010s, Facebook) | — | C |
| Joelle Pineau | FORMER VP of AI Research / head of FAIR - announced departure Apri | RL, reproducibility, health AI | ParlAI/FAIR open science program (2017-2024); Reproducibility checklist (2018) | site | A |
| Johan Schalkwyk | Researcher (joined 2025 from Sesame AI; previously Google) | Speech recognition, voice | Google Voice Search / speech stack (2010s) | — | C |
| Lucas Beyer | Researcher, TBD Lab (joined 2025 from OpenAI Zurich; previously Go | Vision-language pretraining | Vision Transformer / ViT (2020); SigLIP (2023) | — | C |
| Matt Deitke | Researcher (joined 2025 from Allen Institute for AI) | Multimodal models, embodied AI | Molmo (2024); ObjaverseXL (2023) | — | C |
| Nat Friedman | Head of Products & Applied Research, MSL (co-lead of MSL with Wang | Consumer AI products | GitHub (2018-2021, CEO); NFDG venture fund (2022) | site | A |
| Pei Sun | Researcher, TBD Lab (joined 2025 from Google DeepMind/Waymo) | Model training, perception | Waymo Open Dataset (2020) | — | C |
| Rishabh Agarwal | FORMER researcher, TBD Lab - joined 2025, departed ~2025-10 for Pe | RL, distillation, reasoning | Generalized Policy Improvement / on-policy distillation work (2023-2024) | — | C |
| Rob Fergus | VP of AI Research; head of FAIR (appointed 2025-05-08) | Fundamental research leadership | ZFNet (2013); Deconvolutional visualisation of CNNs (2014) | site | A |
| Ruoming Pang | Distinguished/senior researcher, TBD Lab (joined 2025-07 from Appl | Foundation model pretraining | Apple Intelligence Foundation Models (2024); GShard (2020) | — | C |
| Santosh Janardhan | Head of Global Infrastructure, Meta; co-lead of Meta Compute (from | Data centres, silicon, infra software | — | B | |
| Shengjia Zhao | Chief Scientist, Meta Superintelligence Labs (since 2025-07) | Frontier model research direction | ChatGPT (2022); GPT-4 (2023) | — | B |
| Shuchao Bi | Researcher, TBD Lab (joined 2025 from OpenAI) | Multimodal post-training, RL | — | C | |
| Trapit Bansal | Researcher, TBD Lab (joined 2025 from OpenAI) | RL for reasoning | OpenAI o1 (2024, contributor) | — | C |
| Xiaohua Zhai | Researcher, TBD Lab (joined 2025 from OpenAI Zurich; previously Go | Vision-language pretraining, scaling | SigLIP (2023); Scaling Vision Transformers (2022) | — | C |
| Yann LeCun | FORMER VP & Chief AI Scientist - departed November 2025 | World models, self-supervised learning | Convolutional networks / LeNet (1989-1998); I-JEPA (2023); V-JEPA 2 (2025) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Surreal / Egocentric Machine Perception & Spatial AI (Project Aria) | Richard Newcombe (VP, Research Science) — Jakob Julian Engel is Director of Research under him | Egocentric sensing, SLAM/VIO, 3D scene understanding, always-on contextual AI on glasses; | Project Aria (2020); Aria Gen 2 + Aria Gen 2 Pilot Dataset (2025) | largest single RL-R research pillar; exact headcount not public | A |
| Codec Avatars Lab (Pittsburgh) | Sofien Bouaziz (Senior Director, Meta — Codec Avatars). Founded and led until 2025 by Yaser Sheikh (VP), who has left. | Photorealistic driveable human avatars for telepresence — face, full body, hair, relightin | Codec Avatars (2019); Gaussian / Relightable Gaussian Codec Avatars (2024-2025) | Pittsburgh site remains staffed (Alexander Richard's page still reads 'Meta Reality Labs, Pittsburgh' in 2026); size not public | A |
| Neuromotor Interfaces (CTRL-labs at Reality Labs) | Thomas Reardon (VP/Director, Neuromotor Interfaces) and Patrick Kaifosh (Research Science Director) | Non-invasive surface-EMG wristband decoding motor-unit activity into computer input; gener | A generic non-invasive neuromotor interface for human-computer interaction, Nature (2025); Meta Neural Band, shipped with Meta Ray-Ban Display (2025) | not public | A |
| Display Systems Research (optics & displays) | Douglas Lanman (Senior Director, Display Systems Research) | Near-eye displays, holographic and varifocal optics, waveguides, passthrough rendering, pe | Holocake / holographic thin VR optics prototypes (2020-2022); Butterscotch Varifocal retinal-resolution prototype (2022) | not public | C |
| Audio and Haptics research | not individually confirmed in this session | Spatial/personalized audio for AR-VR, acoustic simulation, HRTF personalization, hearing a | Project Aria audio-visual capture; Meta Neural Band haptics (2025) | not public | E |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Alexander Richard | AI Research Scientist, Meta Reality Labs, Pittsburgh | Audio-driven avatars, speech-to-face | MeshTalk (2021); Embody 3D (2025); Audio Driven Real-Time Facial Animation for Social Telepresence (2026) | site | A |
| Andrew Berkovich | Hardware/sensor research lead, Project Aria | Image sensors, low-power vision hardware | Aria Gen 2 Pilot Dataset (2025) | — | D |
| Andrew Bosworth | Chief Technology Officer, Meta; head of Reality Labs | Runs all of Reality Labs | site | A | |
| Armen Avetisyan | Research Scientist, Surreal / Project Aria | Structured 3D reconstruction | Scan2CAD (2019); SceneScript (2024) | — | D |
| Carl Ren | Director / senior lead, Project Aria hardware and research platfor | Aria glasses platform | Project Aria (2020); Aria Gen 2 Pilot Dataset (2025) | — | D |
| Chen Cao | Research Scientist, Codec Avatars | Real-time face avatars from phone scans | Authentic Volumetric Avatars from a Phone Scan (2022); LUCAS (2025); FiCA (2026) | — | D |
| Chen Kong | Research Scientist, Project Aria | Aria sensor/dataset pipeline | Aria Gen 2 Pilot Dataset (2025) | — | D |
| Cheng Peng | Research Scientist, Project Aria | Egocentric 3D reconstruction | EFM3D (2024); Aria Gen 2 Pilot Dataset (2025) | — | D |
| Christian Richardt | Research Scientist, Meta Reality Labs | Novel-view synthesis, neural rendering | URAvatar (2024) | site | C |
| Christopher Xie | Research Scientist, Surreal / Project Aria | 3D scene layout, structured scene models | SceneScript (2024); Human-in-the-Loop Local Corrections of 3D Scene Layouts (2025) | — | D |
| Dejan Markovic | Research Scientist, audio for avatars | Spatial audio, audio-visual telepresence | Implicit HRTF Modeling (2022); Embody 3D (2025) | — | D |
| Douglas Lanman | Senior Director, Display Systems Research | Near-eye displays, computational optics | Holocake holographic optics (2022); Butterscotch Varifocal (2022) | — | C |
| Evonne Ng | Research Scientist, Codec Avatars | Conversational motion and gesture generation | Learning to Listen (2022); From Audio to Photoreal Embodiment (2024); Embody 3D (2025) | — | D |
| Fabian Prada | Research Scientist, Codec Avatars | Clothing/body capture and rendering | Relightable Full-Body Gaussian Codec Avatars (2025); Embody 3D (2025) | — | D |
| Giljoo Nam | Research Scientist, Codec Avatars | Hair capture, appearance modeling | Strand-accurate Multi-view Hair Capture (2019); Gaussian Pixel Codec Avatars (2025); Large-scale Codec Avatars (2026) | — | D |
| Jakob Julian Engel | Director of Research, Meta Reality Labs — egocentric machine perce | SLAM, spatial AI, Aria machine perception | LSD-SLAM (2014); Direct Sparse Odometry / DSO (2016); SceneScript (2024) | site | A |
| Jason Saragih | Research Scientist / Research lead, Codec Avatars | Driveable photoreal face avatars | Deep Appearance Models for Face Rendering (2018); Pixel Codec Avatars (2021); Relightable Full-Body Gaussian Codec Avatars (2025) | — | D |
| Julian Straub | Research Scientist, Surreal / Project Aria | 3D scene understanding, mapping | SceneScript-related layout infilling (2025); Aria Gen 2 Pilot Dataset (2025) | — | D |
| Julieta Martinez | Research Scientist, Reality Labs Research (joined 2022) | 3D human pose, avatar pretraining | A simple yet effective baseline for 3d human pose estimation (2017); Large-scale Codec Avatars (2026) | site | A |
| Junxuan Li | Research Scientist, Codec Avatars | Relightable neural avatars | URAvatar (2024); Relightable Full-Body Gaussian Codec Avatars (2025); Large-scale Codec Avatars (2026) | — | D |
| Kristen Grauman | Research Director, Meta FAIR (NOT RL-R) — UT Austin professor | Egocentric video understanding | Ego4D (2022); Ego-Exo4D (2024) | org · site | A |
| Michael Abrash | Chief Scientist, Reality Labs | Sets RL-R research agenda | Reality Labs Research long-range AR/VR research agenda (2016-2025 Connect keynotes) | org | A |
| Michael Zollhoefer (Zollhöfer) | Research Scientist / Director-level lead, Codec Avatars | Neural rendering, 3D reconstruction | Deep Video Portraits (2018); State of the Art on Neural Rendering (2020); Embody 3D (2025) | — | D |
| Mingfei Yan | Senior lead, Reality Labs Research (Aria / machine perception) | Egocentric perception, on-device models | Aria Gen 2 Pilot Dataset (2025) | — | D |
| Patrick Kaifosh | Research Science Director, Reality Labs Research (CTRL-labs at Rea | sEMG decoding, neuromotor modelling | A generic non-invasive neuromotor interface, Nature (2025) | — | A |
| Rawal Khirodkar | Research Scientist, Meta | Multi-person 3D human perception | Sapiens (2024); URAvatar (2024); Large-scale Codec Avatars (2026) | site | A |
| Richard Newcombe | VP, Research Science, Reality Labs Research (leads the Surreal tea | Egocentric perception, SLAM, spatial AI | KinectFusion (2011); DTAM (2011); Project Aria (2020) | org · site | A |
| Shunsuke Saito | Research Scientist, Meta Codec Avatars Lab | Neural human reconstruction, relightable avatars | PIFu (2019); PIFuHD (2020); Relightable Gaussian Codec Avatars (2024) | site | A |
| Sofien Bouaziz | Senior Director, Meta — leads the Codec Avatars organization | Photoreal avatars, 3D face/body modeling | GenCA (2024); Large-scale Codec Avatars (2026) | site | A |
| Steven Krenn | Research Engineer / capture lead, Codec Avatars | Multi-view capture systems | Embody 3D (2025); Goliath / full-body capture datasets (2024) | — | D |
| Thomas R. Reardon | VP / Director, Neuromotor Interfaces (co-founder & former CEO of C | sEMG neural interfaces | A generic non-invasive neuromotor interface, Nature (2025); Meta Neural Band (2025) | — | A |
| Timur Bagautdinov | Research Scientist, Codec Avatars | Neural avatar representations | Driving-Signal Aware Full-Body Avatars (2021); Relightable Full-Body Gaussian Codec Avatars (2025); FiCA (2026) | — | D |
| Tomas Simon | Research Scientist, Reality Labs Research (Pittsburgh) | Hand/body capture, neural avatars | Hand Keypoint Detection in Single Images using Multiview Bootstrapping (2017); OpenPose (2017); Gaussian Pixel Codec Avatars (2025) | site | D |
| Vasileios Balntas | Research Scientist / manager, Surreal | Local features, relocalization, scene understanding | HPatches (2017); SceneScript (2024) | — | D |
| Vasu Agrawal | Research Engineer/Scientist, Codec Avatars | Capture infrastructure, full-body avatars | Relightable Full-Body Gaussian Codec Avatars (2025); Embody 3D (2025) | — | D |
| Xiaqing Pan | Research Scientist, Project Aria | Aria datasets, digital twins | Aria Digital Twin (2023); Aria Gen 2 Pilot Dataset (2025) | — | D |
| Yaser Sheikh | FORMER VP, Meta (2015-2025); founder of the Meta Reality Lab in Pi | Photoreal telepresence (departed) | OpenPose (2017); Panoptic Studio (2015); Codec Avatars (2019) | site | A |
| Group | Lead | Focus | Flagship | Size | Tier |
|---|---|---|---|---|---|
| Apple Foundation Models (AFM) | Zhifeng Chen | Pre-training, post-training and deployment of the AFM/ADM model family for Apple Intellige | AFM 3 model family: AFM 3 Core / Core Advanced / Cloud / Cloud Pro + ADM 3 Cloud (2026); Apple Intelligence Foundation Language Models, PT-MoE server architecture (2025) | ~100 (reported scale before the 2025 departures; not confirmed by Apple) | B |
| Machine Learning Research (AIML research), publishing arm at machinelearning.apple.com | Samy Bengio (Senior Director of AI and Machine Learning Research) | Open academic publication across the site's 11 declared research areas; conference sponsor | ParaRNN: Large-Scale Nonlinear RNNs, Trainable in Parallel (2026); GSM-Symbolic (2024) and The Illusion of Thinking (2025) | several hundred publishing authors (1,130 papers listed on the site; 459 of them dated 2025-2026) | A |
| Research area: Methods and Algorithms | Largest declared area - optimization, generative modelling, efficiency, learning theory, o | TarFlow / Normalizing Flows are Capable Generative Models (2024); How to Scale Your EMA (2023) | 552 of 1,130 listed papers | A | |
| Research area: Speech and Natural Language Processing | ASR, TTS/expressive voices, LLM post-training, dialogue, Siri language understanding | Siri Expressive Voices / decoupled temporal-depth diffusion audio synthesis (2026); slimIPL and Flashlight/wav2letter ASR line (2021) | 464 papers | A | |
| Research area: Computer Vision | On-device vision backbones, multimodal LLMs, 3D/generative imaging, depth | Ferret / Ferret-UI (2023-2024); Depth Pro (2024) | 298 papers | A | |
| Research area: Human-Computer Interaction | UI understanding and agents, accessibility, interpretability and ML tooling for practition | Talaria: model optimization visualization (2024); Screen Recognition (2021) | 106 papers | A | |
| Research area: Privacy (differential privacy / private learning) | Differential privacy theory and deployment, private federated statistics, memorization | Deep Learning with Differential Privacy lineage (2016); Privacy amplification by shuffling / by iteration (2018-2019) | 73 papers | A | |
| Other declared research areas | Knowledge Bases and Search (62); Data Science and Annotation (62); Tools, Platforms, Frame | MLX (2023); Apple Heart/Health foundation model work (2025) | see counts | A | |
| Internal AIML topic taxonomy (as exposed by the Apple Scholars in AIML PhD fellowship) | The 2026 Apple Scholars cohort is sorted into Apple's own internal research-area names, a | ~19 scholars in the 2026 cohort; the program runs annually since 2020 | A | ||
| Siri / Apple Intelligence engineering (NOT part of AIML research) | Mike Rockwell (VP) | Siri product engineering; moved out of Giannandrea's org in March 2025 and under Craig Fed | Rebuilt Siri shipping on AFM 3 (2026) | unknown | B |
| AKI - Answers, Knowledge and Information (AI web search) | vacant/unclear after Ke Yang's departure; work reported to be gravitating to Eddy Cue's Services org | ChatGPT-like AI answer engine and world knowledge for Siri/Spotlight/Safari | unknown | C |
| Name | Title | Focus | Notable work | Links | Tier |
|---|---|---|---|---|---|
| Afshin Dehghan | Senior manager, Computer Vision | vision for Apple Intelligence | MODUS (2026); VideoFlexTok (2026) | — | D |
| Amar Subramanya | Vice President of AI | runs Apple's AI org | — | A | |
| Benoit Dupin | Senior Director of Machine Learning and AI | ML infrastructure, search partnerships | — | C | |
| Craig Federighi | Senior Vice President, Software Engineering | owns AI and Siri | org | A | |
| Dominik Moritz | Research scientist, data visualization | visualization, model analysis tools | Vega-Lite (2017); Draco (2019); Talaria (2024) | site | D |
| Eddy Cue | Senior Vice President, Services (and Health) | search deals, Google Gemini partnership | org | A | |
| Emmanuel Abbe | Research scientist at Apple (also professor at EPFL) | learning theory, generalization | Exact recovery in the stochastic block model (2016); Generalization on the Unseen (2023) | site | D |
| Fred Hohman | Research scientist, HCI / ML interpretability and tooling | interactive ML tooling, model viz | Summit (2019); Talaria (2024) | site | D |
| Hilal Asi | Research scientist (private optimization) | differentially private optimization | Private Adaptive Gradient Methods (2021) | site | D |
| Iman Mirzadeh | Research scientist (LLM reasoning) | reasoning evaluation, distillation | Teacher Assistant Knowledge Distillation (2020); GSM-Symbolic (2024); The Illusion of Thinking (2025) | site | D |
| Jeffrey Nichols | Research lead, Human-Computer Interaction | UI understanding, screen agents | Screen Recognition (2021); Ferret-UI (2024) | — | D |
| Jiatao Gu | Research scientist (generative modelling) | diffusion, 3D generation, NAR decoding | Non-Autoregressive Neural Machine Translation (2018); NerfDiff (2023); TarFlow (2024) | site | D |
| Josh Susskind (Joshua M. Susskind) | Director / senior research lead, Machine Learning Research | generative models, vision, learning | An Attention Free Transformer (2021); GAUDI (2022); TarFlow / Normalizing Flows are Capable Generative Models (2024) | — | D |
| Kim Vorrath | Vice President, Program Management (moved into the Siri/AI group) | shipping discipline for Siri | — | C | |
| Kunal Talwar | Research scientist (differential privacy, optimization) | differential privacy, private learning | Deep Learning with Differential Privacy (2016); Amplification by Shuffling (2019) | site | D |
| Luca Zappella | Director / senior manager, Machine Learning Research | health and behavioural signals ML | ParaRNN (2026); Uncertainty Quantification for LLM Function-Calling (2026) | — | D |
| Marco Cuturi | Research scientist (optimal transport) | optimal transport, generative flows | Sinkhorn Distances (2013); Soft-DTW (2017); OTT-JAX (2022) | site | D |
| Mehrdad Farajtabar | Research scientist (LLM reasoning, efficiency) | LLM reasoning limits, on-device LLMs | LLM in a flash (2023); GSM-Symbolic (2024); The Illusion of Thinking (2025) | site | D |
| Mike Rockwell | Vice President; leads Siri (and Vision Products Group) | Siri engineering leadership | Apple Vision Pro (2024) | — | B |
| Minsik Cho | Research scientist (on-device efficiency) | LLM compression and on-device inference | LLM in a flash (2023); eDKM (2023) | — | D |
| Navdeep Jaitly | Research scientist (sequence models, speech) | sequence modelling, speech, biology | Pointer Networks (2015); Adversarial Autoencoders (2015) | — | D |
| Oncel Tuzel | Senior research lead, Computer Vision | on-device vision, 3D, efficiency | Region Covariance (2006); MobileOne (2023) | — | D |
| Parikshit Gopalan | Research scientist (theory, calibration) | omniprediction, multicalibration | Omnipredictors (2021); Locally Repairable Codes (2012) | site | D |
| Peter Grasch | Senior engineering leader, Siri / Apple Intelligence | Siri modelling and product ML | MM1.5 (2025); DeepMMSearch-R1 (2026) | — | D |
| Pierre Ablin | Research scientist (optimization) | optimization, training dynamics | Picard / faster ICA by preconditioning (2018); How to Scale Your EMA (2023) | site | D |
| Sabih Khan | Chief Operating Officer | AI infrastructure and compute | org | A | |
| Samy Bengio | Senior Director of AI and Machine Learning Research | runs Apple's academic research arm | Torch (2002); Scheduled Sampling (2015); Understanding deep learning requires rethinking generalization (2017) | site | C |
| Sinead Williamson | Research scientist (Bayesian ML, uncertainty) | Bayesian nonparametrics, uncertainty | Trained on Tokens, Calibrated on Concepts (2026); Uncertainty Quantification for LLM Function-Calling (2026) | site | D |
| Tatiana Likhomanenko | Research scientist (speech recognition) | ASR, semi-supervised speech | wav2letter++ / Flashlight (2018); slimIPL (2021) | site | D |
| Vitaly Feldman | Research scientist (privacy and learning theory) | memorization, private optimization | Does Learning Require Memorization? (2020); Privacy Amplification by Iteration (2018) | site | D |
| Yinfei Yang | Research lead, multimodal / vision-language | multimodal foundation models | MM1 (2024); Universal Sentence Encoder (2018); ALIGN (2021) | site | D |
| Yizhe Zhang | Research scientist (NLP / LLMs) | text generation, LLM post-training | DialoGPT (2020) | site | D |
| Zhe Gan | Research scientist / manager, multimodal foundation models | multimodal LLMs, vision-language | MM1 (2024); Ferret (2023); VILLA (2020) | site | D |
| Zhifeng Chen | Leads the Apple Foundation Models team (senior distinguished engin | foundation model training at scale | TensorFlow (2016); GPipe (2019); GShard (2020) | — | B |