A large-scale community-contributed robotics dataset for vision-language-action learning, featuring 119 datasets from 52 contributors worldwide. This is a converted and curated version of the original HuggingFaceVLA/community_dataset_v1, upgraded to LeRobot v3.0 format.
This dataset was used to pretrain SmolVLA. It was filtered using specific criteria including fps, minimum number of episodes, and qualitative assessment of video quality, using the… See the full description on the dataset page:
https://huggingface.co/datasets/azaracla/community_dataset_v1.