vlaa-thinking-sft dataset transformed to match multimodal-open-r1-8k-verified format with filtering
This dataset was processed using the data-preproc package for vision-language model training.
Base Model: allenai/Molmo-7B-O-0924
Tokenizer: allenai/Molmo-7B-O-0924
Sequence Length: 8192
Processing Type: Vision Language (VL)
input_ids: Tokenized input… See the full description on the dataset page:
https://huggingface.co/datasets/penfever/vlaa-thinking-sft-to-r1-format.