🚧 Coming soon. This model is being optimized and validated for Renesas hardware — no fixed release date yet.
Introduction
GR00T is NVIDIA's generalist foundation model for humanoid robots — a vision-language-action (VLA) model that maps camera input and language instructions to robot actions. Renesas is preparing an optimized deployment of this model for the R-Car Gen5 platform.
Model Architecture: Vision-language-action (VLA) model for robot control.
This page is a placeholder. Content — including the license and source-model link — is provisional and subject to change before the model is fully published.