datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VISION_LANGUAGEA key question for understanding multimodal vs. language capabilities of models is what is
the relative strength of the spatial reasoning and understanding in each modality, as spatial understanding is
expected to be a strength for multimodality? To test this we created a procedurally generatable, synthetic dataset
to testing spatial reasoning, navigation, and counting. These datasets are challenging and by
being procedurally generated new versions can easily be created to combat the effects… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/VISION_LANGUAGE.franka_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "franka",
"total_episodes": 200,
"total_frames": 40394,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:200"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/a3124371940/franka_pnp_language.Marathi-Sign-Language
Marathi Sign Language Detection Dataset Card
Dataset Description
The Marathi Sign Language Dataset is a comprehensive collection of images designed to facilitate the development and training of machine learning models for recognizing Marathi sign language gestures. This dataset includes 43 distinct classes, each representing a unique character in the Marathi sign language alphabet. With approximately 1.2k images per class, the dataset totals over 51k images, all uniformly… See the full description on the dataset page: https://huggingface.co/datasets/VinayHajare/Marathi-Sign-Language.datawhale_eai_pnp_language
Datawhale Every Embodied MuJoCo PnP Language Dataset
This is a small LeRobot-format teaching dataset for the Every Embodied MuJoCo
pick-and-place tutorial. It is used by the Pi0 and SmolVLA examples in:
06-策略抓取或抓取VLA/大模型控制、VLA、VLM/04mujoco复现ACT、Pi0、SmolVLA
Dataset Summary
Format: LeRobot dataset v2.1
Environment: MuJoCo OMY pick-and-place
Episodes: 20
Frames: 2621
FPS: 20
Tasks:
Place the blue mug on the plate.
Place the red mug on the plate.
Main features:… See the full description on the dataset page: https://huggingface.co/datasets/Datawhale/datawhale_eai_pnp_language.omy_pnp_languageomy_pnp_languagerotated-90-demos-languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 90,
"total_frames": 21332,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:90"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/rotated-90-demos-language.omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "omy",
"total_episodes": 20,
"total_frames": 2808,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/HyeonCheol1205/omy_pnp_language.omy_pnp_languageomy_pnp_languagesign_language_image_datasetomy_pnp_languageasl_sign_languages_alphabets_v03omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "omy",
"total_episodes": 20,
"total_frames": 2621,
"total_tasks": 2,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4"… See the full description on the dataset page: https://huggingface.co/datasets/caspar/omy_pnp_language.karsl-502-arabic-sign-language-finaldataset_language_20docs_on_several_languages
Dataset Card for "docs_on_several_languages"
This dataset is a collection of different images in different languages.
The daset includes the following languages: Azerbaijani (az: 0), Belorussian (be: 1), Chinese (zh: 16), English (en: 2), Estonian (et: 3), Finnish (fn: 4), Georgian (gr: 5), Japanese (ja: 6), Korean (ko: 7), Kazakh (kk: 8), Latvian (lv: 10), Lithuanian (lt: 9), Mongolian (mn: 11), Norwegian (no: 12), Polish (pl: 13), Russian (ru: 14), Ukranian (uk: 15).
Each language… See the full description on the dataset page: https://huggingface.co/datasets/AlekseyScorpi/docs_on_several_languages.Indian_sign_language_dataset
Dataset Card for "Indian_sign_language_dataset"
More Information needed
omy_pnp_languagesign_language_image_datasetIndian_sign_language_dataset
Dataset Card for "Indian_sign_language_dataset"
More Information needed
asl_sign_languages_alphabets_v02Indian_Sign_Language_datasetpick_and_place_language_for_train_no_tactile_48This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "Unitree_G1_Inspire",
"total_episodes": 995,
"total_frames": 311048,
"total_tasks": 12,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:995"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4"… See the full description on the dataset page: https://huggingface.co/datasets/eunjuri/pick_and_place_language_for_train_no_tactile_48.omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "omy",
"total_episodes": 11,
"total_frames": 3062,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:11"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sktm502/omy_pnp_language.sign-language-sokdr
Dataset Card for sign-language-sokdr
** The original COCO dataset is stored at dataset.tar.gz**
Dataset Summary
sign-language-sokdr
Supported Tasks and Leaderboards
object-detection: The dataset can be used to train a model for Object Detection.
Languages
English
Dataset Structure
Data Instances
A data point comprises an image and its object annotations.
{
'image_id': 15,
'image': <PIL.JpegImagePlugin.JpegImageFile image… See the full description on the dataset page: https://huggingface.co/datasets/Francesco/sign-language-sokdr.omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "omy",
"total_episodes": 20,
"total_frames": 4247,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/JJinsup/omy_pnp_language.omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "omy",
"total_episodes": 20,
"total_frames": 6226,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Lesoyang/omy_pnp_language.datacomp_small_with_language
Dataset Card for "datacomp_small_with_language"
More Information needed
omy_pnp_languageThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "omy",
"total_episodes": 4,
"total_frames": 709,
"total_tasks": 2,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 20,
"splits": {
"train": "0:4"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Yuri66666/omy_pnp_language.
