160 video rollouts (10 scenes × 16 samples) generated by Wan2.2-TI2V-5B + merged Vidar LoRA
on the RoboTwin put_object_cabinet task. These are the pre-NFT-training (step-0) baseline
samples used to evaluate reward-model behaviour and seed RL fine-tuning.
Base model : Wan2.2-TI2V-5B
LoRA : vidar/merged_vidar_lora.pt (vidar baseline merged into DiT)
Sampler : deterministic ODE (Euler… See the full description on the dataset page:
https://huggingface.co/datasets/VincentNi/wan22-rollout-put-object-cabinet-step0.