This repository contains the data presented in Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT Reasoning.
Project page:
https://sun-hailong.github.io/projects/TVC
Code:
https://github.com/sun-hailong/TVC
A mixture of 345K multimodal long-chain reasoning data.
For more statistics of the dataset, please refer to our paper (coming soon)
LLaVA-OneVision:… See the full description on the dataset page:
https://huggingface.co/datasets/Allen8/TVC-Data.