datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
artelingo-dummyArtELingo is a benchmark and dataset introduced in a research paper aimed at promoting work on diversity across languages and cultures. It is an extension of ArtEmis, which is a collection of 80,000 artworks from WikiArt with 450,000 emotion labels and English-only captions. ArtELingo expands this dataset by adding 790,000 annotations in Arabic and Chinese. The purpose of these additional annotations is to evaluate the performance of "cultural-transfer" in AI systems.
The dataset in ArtELingo… See the full description on the dataset page: https://huggingface.co/datasets/youssef101/artelingo-dummy.dubai-taxi-advertising-synthetic
Dubai Taxi Advertising Compliance (Synthetic)
63 synthetic photorealistic images of Dubai RTA taxis carrying advertising, built
to evaluate whether vision-language models can judge out-of-home (OOH)
advertising compliance rules from a single photograph.
Each image is generated to be an unambiguous pass or fail against one
specific rule from a Dubai taxi advertising technical checklist. The dataset is
an evaluation set — it is small, adversarially balanced, and deliberately… See the full description on the dataset page: https://huggingface.co/datasets/AbdullahKhanSherwani/dubai-taxi-advertising-synthetic.spl-raw-datayolo-rubber-ducks
Rubber Duck Detection Dataset
Overview
This dataset contains 192 annotated images of rubber ducks, specifically curated for object detection tasks. It was used for experimentation related to the YOLOv8n Rubber Duck Detector model.
NOTE: I DO NOT RECOMMEND USING THIS DATASET AT THIS TIME. There is an open and ongoing discussion around the use of the datasets that were combined for this.See related licensing discussion on the forum
Dataset Description… See the full description on the dataset page: https://huggingface.co/datasets/phrogzx/yolo-rubber-ducks.DUALDISCO-perception
Roles
Roles: perception view of DUALDISCO — annot is the source label (conduction / keyhole), kept machine-parseable as the gold for verification and reward parsing; the model reads query + image, where the repo ships microphone sound from a dual-laser powder-bed fusion build in four image encodings as four equal-sized configs — reshaped (consecutive samples arranged as the rows of a grayscale square), scalogram (a continuous-wavelet time-scale view), spectrogram (a short-time… See the full description on the dataset page: https://huggingface.co/datasets/AI4Manufacturing/DUALDISCO-perception.DUALDISCO
Roles
Roles: canon repo — annot is the source label, kept machine-parseable as the gold for verification and reward parsing; there is no filled reasoning column and this repo is not itself a training view. Derived repos each state their own regime on their own card.
DUAL DISCO — is any laser in keyhole mode, from the sound in the air (reasoning track)
Part of the AI4Manufacturing FORGE corpus (Category C, task T-C1), and the corpus's first airborne-acoustic and… See the full description on the dataset page: https://huggingface.co/datasets/AI4Manufacturing/DUALDISCO.dustbin-or-not-dataset
Dustbin vs. Not Dustbin Image Dataset
Dataset Summary
A small binary image classification dataset for predicting whether an image contains a dustbin/trash can.
has_dustbin = 1: dustbin visible
has_dustbin = 0: no dustbin visible
30+ original photographs
Images resized to 224 × 224 RGB
Data Collection
Original photographs were captured by the author and organized into dustbin/ and not_dustbin/ folders. No people, faces, or sensitive personal… See the full description on the dataset page: https://huggingface.co/datasets/srivathsanb14/dustbin-or-not-dataset.dermacheck-temporal-pairs
DermaCheck Temporal Pairs Dataset
Dataset Description
Synthetic temporal image pairs for training MedGemma to detect changes in dermatoscopic images over time.
Created for: MedGemma Impact Challenge 2026 - Novel Task Prize (temporal change detection)
Dataset Statistics
Total pairs: 900
Train: 630 pairs (70.0%)
Validation: 135 pairs (15.0%)
Test: 135 pairs (15.0%)
Generation Methods
Controlled Augmentation (~50%): Original HAM10000 images augmented… See the full description on the dataset page: https://huggingface.co/datasets/dunktra/dermacheck-temporal-pairs.Durian-Plantation-Disease-Leaf-Rot-Detection-Image-Dataset
Durian Plantation Disease Leaf Rot Detection Image Dataset
Ensures high-quality data through multiple rounds of annotation and automated consistency checks, combined with reviews by agricultural pathology experts. The annotation team consists of 10 professionals in agriculture and computer vision. Pre-processing steps include noise reduction, size adjustments, and color normalization to enhance the model's recognition capabilities. Data is stored in JPG format, organized… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Durian-Plantation-Disease-Leaf-Rot-Detection-Image-Dataset.VQA-BronchoscopyBM-UET
VinDr-CXR-VQA Dataset
Dataset Description
VinDr-CXR-VQA is a large-scale chest X-ray Visual Question Answering (VQA) dataset designed for explainable medical AI with spatial grounding capabilities. The dataset combines natural language question-answer pairs with bounding box annotations and clinical reasoning explanations.
Key Features
🏥 4,394 chest X-ray images from VinDr-CXR
💬 17,597 question-answer pairs across 6 question types
📍 Spatial grounding with… See the full description on the dataset page: https://huggingface.co/datasets/DungNgoc/VQA-BronchoscopyBM-UET.Durian-Plantation-Epiphyte-Moss-Image-Dataset
Durian Plantation Epiphyte Moss Image Dataset
Current durian plantations face the challenge of identifying epiphytic moss, which affects plant health and yield. Existing solutions often rely on manual monitoring, which is time-consuming, labor-intensive, and lacks accuracy. This dataset aims to improve the detection efficiency and accuracy of epiphytic moss through automated image recognition technology. Data collection was carried out using drones and high-definition cameras in… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Durian-Plantation-Epiphyte-Moss-Image-Dataset.Durian-Plantation-Disease-Leaf-Rot-Detection-Image-Dataset
Durian Plantation Disease Leaf Rot Detection Image Dataset
Ensures high-quality data through multiple rounds of annotation and automated consistency checks, combined with reviews by agricultural pathology experts. The annotation team consists of 10 professionals in agriculture and computer vision. Pre-processing steps include noise reduction, size adjustments, and color normalization to enhance the model's recognition capabilities. Data is stored in JPG format, organized… See the full description on the dataset page: https://huggingface.co/datasets/shangzx/Durian-Plantation-Disease-Leaf-Rot-Detection-Image-Dataset.cv_backbones_duplicate
Dataset Card for "monetjoe/cv_backbones"
This repository consolidates the collection of backbone networks for pre-trained computer vision models available on the PyTorch official website. It mainly includes various Convolutional Neural Networks (CNNs) and Vision Transformer models pre-trained on the ImageNet1K dataset. The entire collection is divided into two subsets, V1 and V2, encompassing multiple classic and advanced versions of visual models. These pre-trained backbone… See the full description on the dataset page: https://huggingface.co/datasets/EF54321/cv_backbones_duplicate.DualStream-Foundational-Manifests
Dual-Stream DeepFake Foundational Baseline Connectors
This repository provides standardized data connectors, download manifests, and partition splits for the 8 foundational baseline datasets used in the Dual-Stream Deepfake Detection Framework.
📊 Dual-Stream Model Allocation
🖼️ Model 1: General Vision & Signal Model (>578,000 samples)
NTIRE-RobustAIGenDetection (~120,000 samples): Multi-generator synthetic artifacts (MSU 2024).
CIFAKE (120,000… See the full description on the dataset page: https://huggingface.co/datasets/ThangCao/DualStream-Foundational-Manifests.robot-vision-sample-dataset12
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
sonic-forage-strawberry-logo-dude-22
Sonic Forage Strawberry Logo Dude 22
A 22-image public-safe generated reference dataset for Sonic Forage, positioned as the World’s first autonomous DJ operating system.
Each sample depicts a strawberry AGI DJ mascot / logo dude with a friendly third eye, headphones, decks, and different renegade-set environments. The generation prompts required: no text, no words, no letters, no logos, no watermark, no drug-use depiction, no alcohol, no weapons, and no real-person likeness.… See the full description on the dataset page: https://huggingface.co/datasets/Sonic-Forage/sonic-forage-strawberry-logo-dude-22.robot-vision-dataset-vk
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-vx
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-sample-dataset-vv
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-vl
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-sample-dataset
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-sample-dataset11
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-sample-dataset31
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot_vision_sample_dataset_a12
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-model-vs
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-vf
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-vg
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot_vision_dataset_vh
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-vj
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
robot-vision-dataset-v2
Robot Vision Sample Dataset
This dataset is created for experimental robotics vision and perception training purposes.
