Views
No views yet
RIPT-VLA enables interactive post-training for any pretrained Vision-Language-Action (VLA) model using only sparse binary success rewards.
With K-rollout interaction, dynamic sampling, and leave-one-out advantage estimation, RIPT-VLA achieves state-of-the-art performance in extremely low-data regimes.
| Suite | SFT Checkpoint | RIPT Checkpoint |
|---|---|---|
| LIBERO-90 | ✅ | ✅ |
| LIBERO-GOAL | ✅ | ✅ |
| LIBERO-LONG | ✅ | ✅ |
| LIBERO-OBJECT | ✅ | ✅ |
| LIBERO-SPATIAL | ✅ | ✅ |
| Suite | SFT Scale Head | RIPT LoRA Adaptor |
|---|---|---|
| LIBERO-GOAL | ✅ | ✅ |
| LIBERO-LONG | ✅ | ✅ |
| LIBERO-OBJECT | ✅ | ✅ |
| LIBERO-SPATIAL | ✅ | ✅ |