This repository contains the data presented in Video-R1: Reinforcing Video Reasoning in MLLMs.
Code:
https://github.com/tulerfeng/Video-R1
Video data folder: CLEVRER, LLaVA-Video-178K, NeXT-QA, PerceptionTest, STAR
Image data folder: Chart, General, Knowledge, Math, OCR, Spatial
Video-R1-COT-165k.json is for SFT cold start, and Video-R1-260k.json is for RL training.
Data Format in Video-R1-COT-165k:
{
"problem_id": 2,
"problem": "What appears on the screen in Russian during the… See the full description on the dataset page:
https://huggingface.co/datasets/Huangzx1023/Video-Training.