datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
record-test1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 4,
"total_frames": 2990,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:4"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gomipapa/record-test1.lekiwi_gomaxlekiwi_gomax2NYCU_Cup_Stackingdataset_v3_synth_top50custom_selena_gomez_left_2dataset_v3_synth_top50custom_selena_gomez_middle_2cafe_stretch_groot_idm_labeled
Cafe Stretch — GR00T v2 corpus (mixed GT / IDM labels)
Hello Robot Stretch serve task (cup → coffee machine), 656 episodes / 1,180,922 frames @ 6 fps,
in GR00T v2 (LeRobot-compatible) format. Labels are mixed by domain — the exact recipe we use
for policy training:
Original episodes carry ground-truth (GT) actions (real teleop commands, full amplitude).
Augmented episodes carry IDM-inferred actions — an Inverse Dynamics Model trained on the GT
episodes re-labels the visually… See the full description on the dataset page: https://huggingface.co/datasets/Gom-sy/cafe_stretch_groot_idm_labeled.dataset_v3_synth_top50custom_selena_gomez_right_2
