Views
No views yet
so101-sim-pickplace-v2
(160 SO-ARM100 episodes incl. 12.5% verified-retry recovery demonstrations):"Pick up the red block and place it in the blue tray."

| Policy | Success | |
|---|---|---|
| SmolVLA base, zero-shot | 0% | |
| SmolVLA on 100 nominal eps | 55% | all failures = right-side non-engagements (coverage) |
SmolVLA on the 160-ep -v2 set (this model) | 90% | coverage fixed; one eval episode shows a live miss→retry→success (165 steps vs the stereotyped ~105) |
ACT (~52M specialist BC) on the same -v2 data | 50% | hurt by the same data — see below |
-v2 dataset moved two architectures in opposite directions:Trained on -v2 (160 eps, incl. recovery branches) | vs. its nominal-data baseline |
|---|---|
| ACT (chunked L1 regression) | 65% → 50% (and temporal ensembling flips from +10 to −15) |
| SmolVLA (flow-matching action expert) | 55% → 90% |
lerobot-train: batch 8, 12k steps (~8 h), AMP, frozen VLM /
action-expert-only (defaults), cameras renamed to the base's naming
(front→camera1, wrist→camera2). Same recipe as the
55% v1 model —
only the data changed.trs_so_arm100).