Views
No views yet
MrBalanceMLPForRL64128 (first, second and third layer)64 (last layer)428512620480.990.950.200.500.0050.503e-41e-5true1275610000| Object | Reward Mean | Len Mean | Survival % | Tracking Err |
|---|---|---|---|---|
| sphere | 24,398.48 | 10,000.0 | 100.0% | 0.0205m |
| egg | 23,240.42 | 10,000.0 | 100.0% | 0.0458m |
| heavy_ball | 24,310.68 | 10,000.0 | 100.0% | 0.0401m |
| Object | Reward Mean | Len Mean | Survival % | Tracking Err |
|---|---|---|---|---|
| sphere | 24,398.48 | 10,000.0 | 100.0% | 0.0205m |
| disk | 21,191.56 | 10,000.0 | 100.0% | 0.1362m |
| egg | 23,240.42 | 10,000.0 | 100.0% | 0.0458m |
| cup | 21,702.17 | 10,000.0 | 100.0% | 0.1242m |
| coin | 20,973.76 | 10,000.0 | 100.0% | 0.1428m |
| stick | 13,601.42 | 8,477.4 | 80.0% | 0.3399m |
| tall | 22,265.77 | 10,000.0 | 100.0% | 0.1077m |
| triangle | 18,794.99 | 10,000.0 | 100.0% | 0.2181m |
| block | 21,489.38 | 10,000.0 | 100.0% | 0.1316m |
| puck | 21,265.33 | 10,000.0 | 100.0% | 0.1361m |
| cone | 18,327.38 | 10,000.0 | 100.0% | 0.2316m |
| capsule | 22,302.92 | 10,000.0 | 100.0% | 0.1097m |
| wedge | 18,921.33 | 10,000.0 | 100.0% | 0.2148m |
| tetra | 18,838.96 | 10,000.0 | 100.0% | 0.2165m |
| flat_bar | 20,550.93 | 10,000.0 | 100.0% | 0.1526m |
| cross | 21,394.63 | 10,000.0 | 100.0% | 0.1435m |
| L_shape | 21,221.41 | 10,000.0 | 100.0% | 0.1389m |
| wide_block | 20,857.84 | 10,000.0 | 100.0% | 0.1611m |
| heavy_ball | 24,310.68 | 10,000.0 | 100.0% | 0.0401m |
| offcenter_block | 21,408.22 | 10,000.0 | 100.0% | 0.1395m |
pip install torch transformers safetensors "mujoco==3.10.0" numpyinference.py and balance_plate_rl.py and run:python inference.py--render for live visual rendering or/and --object to choose a specific object.@misc{mrbalance,
title = {Mr. Balance: Teaching RL agents to Balance Objects},
organization = {FromZero},
authors = {Paul Courneya},
year = {2026},
url = {https://huggingface.co/fromziro/MrBalance]
}