Views
No views yet
push_cube task.| Setting | Value |
|---|---|
| Dataset | taewonkoo/push_cube_v2_rlds |
| Optimizer step | 10,000 / 20,000 |
| Per-device batch size | 4 |
| Gradient accumulation | 4 |
| Effective batch size | 16 |
| Learning rate | 1e-4 |
| Scheduler | Cosine with 2,000 warmup steps |
| LoRA rank | 32 |
| Input images | 2 (front, top) |
| Proprioception | Enabled, 6 dimensions |
| Action dimension | 6 |
| Action chunk length | 8 |
| Precision | BF16 on an RTX 3090 |
model.safetensors: VLA model with the LoRA weights mergedlora_adapter/: trainable LoRA adapteraction_head--10000_checkpoint.pt: continuous action headproprio_projector--10000_checkpoint.pt: proprioception projectorvla_extras--10000_checkpoint.pt: trainable VLA extras such as action queriestraining_state--10000_checkpoint.pt: optimizer, scheduler, scaler, and RNG statemodel.safetensors alone is not the
complete policy.