datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai2thor-perspective-qa-20k-balanced-splits-with-objai2thor-perspective-qa-20k-raw-splitsshelf_bin_long_splitThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "easo",
"total_episodes": 2268,
"total_frames": 744700,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 3,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:2268"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/willx0909/shelf_bin_long_split.anyedit-splitai2thor-perspective-qa-100k-balanced-training-v1-splitsmvtec_all_objects_splitopen-images-v7-subset-splitsnabirds_custom_split_preprocessedsync_bigjob_8_finalised_processed_with_error_handling_from_51th_splitcolpali_train_set_split_by_sourcecoco-caption-splitai2thor-perspective-qa-800-qa-v7-splitsai2thor-perspective-qa-800-qa-v9-splitsof_filtered_splitminc-2500_split_1
Materials in Context Dataset (MINC-2500)
Dataset Summary
(from the website)
MINC-2500 is a patch classification dataset with 2500 samples per category
(Section 5.4 of the paper). This is a subset of MINC where samples have been
sized to 362 x 362 and each category is sampled evenly. The original resolution
images are not needed as we include the extracted patches in the archive.
Recap-DataComp-1B_split_3sec-material-contracts-qa-splittedMixed and filtered version of chenghao/sec-material-contracts-qa and jordyvl/DUDE_subset_100val.
Recap-DataComp-1B_split_4ai2thor-perspective-qa-2000-qa-v5-splitsai2thor-perspective-qa-1000-qa-v5-splitsai2thor-perspective-qa-annotated-411-splitsai2thor-perspective-qa-800-qa-v8-splitscolpali-train-set-splitted-translatedRecap-DataComp-1B_split_5messytable-split-multiviewplace_blue_bottle_splitThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "a2d",
"total_episodes": 325,
"total_frames": 183325,
"total_tasks": 4,
"total_videos": 975,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 25,
"splits": {
"train": "0:325"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Pi-robot/place_blue_bottle_split.verl_format_batched_splits_v1Recap-DataComp-1B_split_7VLLM_ChartQA_splitRecap-DataComp-1B_split_8
