datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AgiBotWorld-Beta_G1_task_504_Supplement_supermarket_snacks
agibot_task_504
This dataset converts the AgiBot format uniformly into LeRobot V3.0.
Dataset Statistics
robot_name: G1
end_effector: 夹爪
task: 补充超市零食part_1
total_episodes: 2355
total_tasks: 1
size: 342G
Dataset Structure
├── data
│ └── chunk-xxx
│ ├── file-xxx.parquet
├── meta
│ ├── episodes
│ │ └── chunk-xxx
│ │ └── file-xxx.parquet
│ ├── info.json
│ ├── stats.json
│ └── tasks.parquet
└── videos
├──… See the full description on the dataset page: https://huggingface.co/datasets/BAAI-DataCube/AgiBotWorld-Beta_G1_task_504_Supplement_supermarket_snacks.Emo-2-SNACparler-tts_mls_eng_10k_snac_token_old
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/blanchon/parler-tts_mls_eng_10k_snac_token_old.AgiBotWorld-Beta_G1_task_506_Supplement_supermarket_snacks
agibot_task_506
This dataset converts the AgiBot format uniformly into LeRobot V3.0.
Dataset Statistics
robot_name: G1
end_effector: 夹爪
task: 补充超市零食part_2
total_episodes: 773
total_tasks: 1
size: 83G
Dataset Structure
├── data
│ └── chunk-xxx
│ ├── file-xxx.parquet
├── meta
│ ├── episodes
│ │ └── chunk-xxx
│ │ └── file-xxx.parquet
│ ├── info.json
│ ├── stats.json
│ └── tasks.parquet
└── videos
├──… See the full description on the dataset page: https://huggingface.co/datasets/BAAI-DataCube/AgiBotWorld-Beta_G1_task_506_Supplement_supermarket_snacks.snac-testTask_99999_pick_place_snack_JM_MCAP
Task_99999_pick_place_snack_JM_MCAP
Created with Cyclo Intelligence by ROBOTIS.
emilia-en-snac
Stats (EN)
Emilia: 46,349 hours
Emilia-YODAS: 87,258 hours
Total: 133,607 hours
License
The Emilia subset is licensed under CC BY-NC 4.0.
The Emilia-YODAS subset is licensed under CC BY 4.0.
Reference
@inproceedings{emilialarge,
author={He, Haorui and Shang, Zengqiang and Wang, Chaoren and Li, Xuyuan and Gu, Yicheng and Hua, Hua and Liu, Liwei and Yang, Chen and Li, Jiaqi and Shi, Peiyang and Wang, Yuancheng and Chen, Kai and Zhang, Pengyuan and Wu… See the full description on the dataset page: https://huggingface.co/datasets/nytopop/emilia-en-snac.egopi_latal_openarm_snacksnac_llm_parler_ttssnacso100_grasp_snackThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100",
"total_episodes": 37,
"total_frames": 16598,
"total_tasks": 1,
"total_videos": 74,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:37"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/masato-ka/so100_grasp_snack.Hypa-Voices-snac
Hypa-Voices-snac
This repository is the SNAC token companion to hypaai/Hypa-Voices.
Both datasets belong to the Hypa-Voices collection and contain the same 8,800 curated records with identical metadata columns. The only difference is how speech is stored:
Repository
Speech field
Format
hypaai/Hypa-Voices
audio
FLAC (decoded waveform)
hypaai/Hypa-Voices-snac (this repo)
codes_list
SNAC discrete token sequence
For the full dataset description, data fields… See the full description on the dataset page: https://huggingface.co/datasets/hypaai/Hypa-Voices-snac.snacks
Dataset Card for Snacks
Dataset Summary
This is a dataset of 20 different types of snack foods that accompanies the book Machine Learning by Tutorials.
The images were taken from the Google Open Images dataset, release 2017_11.
Dataset Structure
Number of images in the train/validation/test splits:
train 4838
val 955
test 952
total 6745
Total images in each category:
apple 350
banana 350
cake 349
candy 349
carrot… See the full description on the dataset page: https://huggingface.co/datasets/Matthijs/snacks.libritts-snac-tokens
libritts-snac-tokens
To learn about Trelis Enterprise Voice Services, see Trelis.com/voice-ai-services.
LibriTTS-R encoded with hubertsiuzdak/snac_24khz (hierarchical RVQ, 3 levels at 12 / 24 / 48 fps, 4,096 entries each).
Orpheus-style interleave per 1/12-sec audio frame: [L0[t], L1[2t], L1[2t+1], L2[4t], L2[4t+1], L2[4t+2], L2[4t+3]]. 7 tokens per audio frame, 84 fps flat.
Offset vocab 12,288: L0 in [0, 4096), L1 in [4096, 8192), L2 in [8192, 12288). Decode with level = token //… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/libritts-snac-tokens.Emilia-All-EN-Snac-LLama3.2emilia-snac-with-spk-embsnackbasue
Bangumi Image Base of Snack Basue
This is the image base of bangumi Snack Basue, we detected 30 characters, 3668 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the characters'… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/snackbasue.emilia-snac-with-spk-emb-ZHemilia-basic-snac-with-spk-embeval_pi05_multi_packing_200k_03012026_snacksThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "bi_yam_follower",
"total_episodes": 50,
"total_frames": 189142,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Jiafei1224/eval_pi05_multi_packing_200k_03012026_snacks.Task_99999_pick_place_snack_JM_MCAP_lerobot_v21_processed
Task_99999_pick_place_snack_JM_MCAP
Created with Cyclo Intelligence by ROBOTIS.
Task_99999_pick_place_snack_JM_MCAP_lerobot_v21
Task_99999_pick_place_snack_JM_MCAP
Created with Cyclo Intelligence by ROBOTIS.
laions_got_talent_orpheus_snacSome Laion's Got Talent (https://huggingface.co/datasets/laion/laions_got_talent) voice snippets converted to snac tokens in the format of the Orpheus-TTS https://github.com/canopyai/Orpheus-TTS
We converted the data into instructions like format.
The snac data is in 7 token frame groups. See the Orpheus blog for more details: https://canopylabs.ai/model-releases
We did not create the original dataset and are only providing snac token with minimal text instructions for ease of use.
You must be… See the full description on the dataset page: https://huggingface.co/datasets/laion/laions_got_talent_orpheus_snac.game-novel-snacTask_49506_SnackBoardMerged_lerobot
dataset-robotis-task-49506-snackboardmerged-lerobot
eval_pi05_multi_packing_200k_03122026_snacks_02This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "bi_yam_follower",
"total_episodes": 25,
"total_frames": 92816,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:25"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cortexairobot/eval_pi05_multi_packing_200k_03122026_snacks_02.omx-pick-snack-100
Task_006_Pick Snack_MCAP
Created with Cyclo Intelligence by ROBOTIS.
snack_feeder_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 30,
"total_frames": 13499,
"total_tasks": 1,
"total_videos": 30,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:30"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/neka-nat/snack_feeder_test.Task_50345612_SnackBottleInterv_lerobot
dataset-temp-task-50345612-snackbottleinterv-lerobot
Task_99999_pick_place_snack_cut_processed
Task_99999_pick_place_snack_cut
Created with Cyclo Intelligence by ROBOTIS.
