datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagefolder_with_metadata_no_splitslagenda_split
LAGENDA Dataset
This is a community mirror of the LAGENDA dataset created by LayerTeam. It has been uploaded here for easier access and integration with the Hugging Face datasets library.
All credit, rights, and accolades belong to the original authors. Please see the citation section below.
Dataset Description
LAGENDA (Large Age and Gender Dataset) is a dataset designed for age and gender recognition tasks. It addresses common biases in existing datasets by ensuring a… See the full description on the dataset page: https://huggingface.co/datasets/uaebn/lagenda_split.ai2thor-perspective-qa-20k-balanced-splits-with-objai2thor-perspective-qa-20k-raw-splitsshelf_bin_long_splitThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "easo",
"total_episodes": 2268,
"total_frames": 744700,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 3,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:2268"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/willx0909/shelf_bin_long_split.anyedit-splitai2thor-perspective-qa-100k-balanced-training-v1-splitsRSCC-RSEdit-Test-Split
RSCC-RSEdit-Test-Split
This directory contains the test split for RSCC-RSEdit dataset.
Directory Structure
RSCC-RSEdit-Test-Split/
├── images/ # Original images (676 PNG files)
├── masks/ # Original grayscale masks (338 PNG files)
│ └── [mask files with pixel values 0,1,2,3,4]
├── masks_colorful/ # Colorful RGBA visualization masks (338 PNG files)
│ └── [same filenames as masks/, but in RGBA format with colors]
├──… See the full description on the dataset page: https://huggingface.co/datasets/BiliSakura/RSCC-RSEdit-Test-Split.Chinese_Children_Image_Captioning_Dataset_Split0
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split0.mvtec_all_objects_splitrplan-evacuation-test-splitopen-images-v7-subset-splitsnabirds_custom_split_preprocessedsync_bigjob_8_finalised_processed_with_error_handling_from_51th_splitcolpali_train_set_split_by_sourcecoco-caption-splitai2thor-perspective-qa-800-qa-v7-splitsai2thor-perspective-qa-800-qa-v9-splitsof_filtered_splitChinese_Children_Image_Captioning_Dataset_Split1
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split1.minc-2500_split_1
Materials in Context Dataset (MINC-2500)
Dataset Summary
(from the website)
MINC-2500 is a patch classification dataset with 2500 samples per category
(Section 5.4 of the paper). This is a subset of MINC where samples have been
sized to 362 x 362 and each category is sampled evenly. The original resolution
images are not needed as we include the extracted patches in the archive.
Recap-DataComp-1B_split_3sec-material-contracts-qa-splittedMixed and filtered version of chenghao/sec-material-contracts-qa and jordyvl/DUDE_subset_100val.
Recap-DataComp-1B_split_4ai2thor-perspective-qa-2000-qa-v5-splitsUNO1m-filtered-splitai2thor-perspective-qa-1000-qa-v5-splitsai2thor-perspective-qa-annotated-411-splitsai2thor-perspective-qa-800-qa-v8-splitsnser-ibvs-mask-splitter-dataset
NSER-IBVS Mask Splitter Dataset
Dataset Description
This dataset is used to train the Mask Splitter neural network, a key component of the NSER-IBVS visual servoing
framework for autonomous drone control. The network learns to split a vehicle segmentation mask into front
and back regions, enabling the analytical IBVS controller to compute precise velocity commands.
Associated Resources
Resource
Link
Paper
ICCV 2025 Workshop
arXiv… See the full description on the dataset page: https://huggingface.co/datasets/brittleru/nser-ibvs-mask-splitter-dataset.
