[Project Page] | [Paper] | [Code]
This repository provides datasets collected with the UMI, converted into the LeRobot data format, along with synthetic vision-language data used in the paper OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning.
The robot data covers two main tasks:
cocktailContains 299 real-world demonstrations collected in the lab, each with reasoning… See the full description on the dataset page:
https://huggingface.co/datasets/ad1t7a/onetwovla-dataset.