This repo contains a full experiment log for fine-tuning a pi0_fast_piper_lora policy on a Piper pick-and-place dataset, then validating the policy on held-out episodes with side-by-side MuJoCo playback:
Left video: model-predicted actions
Right video: ground-truth actions