Fine-tuned pi0.5 checkpoint for SO101 magnetic-cube stacking. Trained on Isambard GH200 (4× GPU node). Training is in progress — this repo publishes interim checkpoints under step_<N>/params/; final 50k may be added as step_49999/params/.
Published checkpoints
Directory
Step
Loss (approx)
Notes
step_10000/params/
10,000
0.0179
Interim; root MODEL_PASSPORT.json / SIGNOFF.json
step_15000/params/
15,000
0.0151
Interim; passport/signoff under step_15000/
Each step directory may include its own MODEL_PASSPORT.json and SIGNOFF.json when published separately from the root (10k) package.
Experiment
Objective: Fine-tune pi0.5 on lorenzouttini/so101_stacking_magnetic_cubes.
Weight init: pi0.5 base (pi05_base).
Training target: 50,000 steps (save_interval=5000).