Glossary · term
GR00T N1.6
A reference generalist VLA model for humanoid robots (NVIDIA): a VLM backbone (a Cosmos-2B variant) plus a 32-layer diffusion transformer, trained on thousands of hours of teleoperation data across multiple robot bodies. It embodies the 2025-2026 consensus: a pretrained VLM plus an action-generation module. Showcased by NVIDIA at CES 2026.
Training2026Wave 3 · 2025–26Maturity: 3/5
Maturity rationale
NVIDIA Isaac humanoid foundation model
References
Author: Jensen Huang