World-R1 is a prompt-only dataset for text-to-video world simulation. It accompanies World-R1: Reinforcing 3D Constraints for Text-to-Video Generation, where reinforcement learning is used to improve 3D consistency while preserving visual quality and motion diversity in generated videos.The dataset contains English prompts that describe static environments, dynamic scenes, and camera-aware video generation scenarios. It is designed for post-training… See the full description on the dataset page:
https://huggingface.co/datasets/microsoft/World-R1.