This dataset was created using LeRobot.
Successful BananaInBowl demonstrations collected from an RFCL-trained SAC policy in
Isaac Lab (RoboLab), for behaviour-cloning research on strategy diversity.
Lane: metra-alpha — a sweep of the METRA intrinsic-reward scale alpha, with every
other axis held fixed (banana task, 50 demos, sf=0.5, z_dim=3, phi_space=full, z_unit=true).
Each dataset is one (difficulty level, alpha) cell. Difficulty here: medium.
Collection:… See the full description on the dataset page:
https://huggingface.co/datasets/DAVIAN-Robotics/metra-alpha-ma-medium-a05-n3000.