[!Note]
This repository contains an expanded version of Vero-600K, using the same dataset curation and filtering process, but with a larger set.
Note that task categories are not balanced in this dataset.
Vero is a fully open reinforcement learning (RL) recipe for training and evaluating multi-task visual reasoning with vision-language models. This repository contains the Vero-600K dataset, a curation of 600K reinforcement learning samples from 59 datasets… See the full description on the dataset page:
https://huggingface.co/datasets/zlab-princeton/Vero-1.6M.