stray-light/zerograv-manipulation-trajectories-v0
⚠️ Superseded — do not use for distillation These trajectories were collected from policies trained before the sim-real action-space alignment. Measured on this data: mean |arm action| is 0.72, with 46-53% of actions at the clip boundary of [-1, 1]. For comparison, the real robot's own pi0.5-DROID policy outputs a mean of 0.066 and human teleop 0.101 — so these actions sit 7-11x outside the distribution the base model has seen. Distilling them would push pi0.5-DROID's action… See the full description on the dataset page: https://huggingface.co/datasets/stray-light/zerograv-manipulation-trajectories-v0.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face