CoolFace
Datasetpublic

Auryal/libero-microwave-grm-rollouts

LIBERO Microwave GRM Rollouts Dense-reward-annotated rollout dataset from GRPO training of OpenVLA-OFT on LIBERO-10 Task 9 ("put the yellow and white mug in the microwave and close it"). Dataset Stat Value Episodes ~2,100 (T >= 5 steps) Format LeRobot (parquet + images) Task put the yellow and white mug in the microwave and close it Size 41 GB Reward model Robo-Dopamine GRM-3B Policy OpenVLA-OFT (LoRA, GRPO-trained) Simulator LIBERO… See the full description on the dataset page: https://huggingface.co/datasets/Auryal/libero-microwave-grm-rollouts.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes7.5kdownloads

Auryal/libero-microwave-grm-rollouts · main · files are served by the source, never re-hosted here