datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
InsightVQA
InsightVQA: High-Dimensional Emotion-Cognitive Visual Question Answering Benchmark
Overview
InsightVQA is a large-scale dataset designed for hierarchical visual question answering that bridges emotion understanding and cognitive reasoning. While existing benchmarks predominantly focus on surface-level emotion recognition , InsightVQA introduces a structured paradigm to evaluate a model's ability to interpret emotional causes, ground evidence, and reason about… See the full description on the dataset page: https://huggingface.co/datasets/ziyul707/InsightVQA.viet-cultural-vqa
🇻🇳 Vietnamese Cultural VQA Dataset
📖 Dataset Description
The Vietnamese Cultural VQA Dataset is a comprehensive multimodal dataset designed for Visual Question Answering (VQA) tasks focused on Vietnamese cultural heritage. This dataset aims to bridge the gap in understanding and preserving Vietnamese culture through AI-powered visual understanding and question answering.
🎯 Dataset Summary
📊 Total Images: 28,505 high-quality cultural images
💬 Total… See the full description on the dataset page: https://huggingface.co/datasets/IAmFuch/viet-cultural-vqa.UrbanPersona-120K-Interpretive
UrbanPersona-120K-Interpretive
Annotation corpora and analysis outputs for "Persona Prompting in Multimodal Urban Perception: Descriptive Convergence and Interpretive Variation" (EMNLP 2026 Workshop Pandora). Two
multimodal LLMs, Qwen3-VL-8B and Gemma4 E4B, annotate the same 50 PerceptSent urban scenes as the
same 1,200 demographic personas at T = 0.1, 60,000 persona × image attempts per model and 120,000
in all, next to their no-persona ablations, a greedy T = 0 decoding… See the full description on the dataset page: https://huggingface.co/datasets/MInDS-lab-UTFPR/UrbanPersona-120K-Interpretive.pad-auto-solver-reviewed
PAD Reviewed Dataset
Canonical reviewed PAD board/orb artifacts for dw-indie/pad-auto-solver-reviewed. This repository
contains immutable reviewed package revisions and does not contain raw captures,
training runs, checkpoints, or model binaries.
Packages exported: 28
Active catalog datasets: 14
Catalog schema: 3
Layout
packages/<dataset_id>.tar: deterministic self-contained reviewed package
catalog.json: active revision heads and coverage summary… See the full description on the dataset page: https://huggingface.co/datasets/dw-indie/pad-auto-solver-reviewed.rule34lol-images-part2
Dataset Card for rule34lol-images-part2
Dataset Summary
This dataset contains information about image files from rule34.lol, a booru-style imageboard. The dataset includes metadata for 77,000 image files, including URLs, tags, file information, and like counts. The actual image files are stored in zip archives, with each archive containing 1000 image files (except the last archive). This is Part 2 of 2 for the complete rule34lol-images dataset. Part 1 can be found here.… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/rule34lol-images-part2.OpenGameArt-GPL-3.0
Dataset Card for OpenGameArt-GPL-3.0
Dataset Summary
This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the GNU General Public License version 3.0 (GPL-3.0). The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, and textures along with their associated metadata.
Languages
The dataset is primarily monolingual:
English (en): All asset descriptions… See the full description on the dataset page: https://huggingface.co/datasets/irfankabir02/OpenGameArt-GPL-3.0.ikea-us-products-2025
IKEA US Product Dataset (July 2025)
This dataset is a structured snapshot of ~30,000 IKEA US products, scraped from the official IKEA US website in July 2025.
It contains product metadata (titles, descriptions, categories, materials, care instructions, etc.) and associated product images.
Contents
products-us.jsonl — one JSON object per product with structured fields.
images-us/ — the first "hero" image for each product, downloaded via image_downloader_first.py.… See the full description on the dataset page: https://huggingface.co/datasets/jeffreyszhou/ikea-us-products-2025.OpenHotelsSample
OpenHotels Representative Sample
This repository contains a representative sample of OpenHotels for review and inspection. It mirrors the full OpenHotels release structure: image files are stored in tar shards under shards/, and metadata files describe the gallery, non-object query images, object-centric query images, and hotel classes.
The sample is intended for data-quality inspection, not benchmark reporting. Use the full OpenHotels dataset for final evaluation.… See the full description on the dataset page: https://huggingface.co/datasets/imagingforgood/OpenHotelsSample.Latent-Resonance-AI-Image-Forensics-Benchmark-N100
Latent Resonance: SOTA Empirical AI Image Forensics Benchmark (N=100 & N=1,000 Scale)
Author: Debdip Bandyopadhyay (Independent AI Researcher, Kolkata, India; M.Tech, IIT Jodhpur, AI & Data Science)Preprint & Paper: Latent Resonance: Zero-Shot Autoencoder Inversion and Azimuthal Spectral Forensics for Diffusion Image Attribution (IEEE Flagship / CERN Zenodo 2026)
Benchmark Overview
This repository provides:
The official verified $N=100$ ground-truth image… See the full description on the dataset page: https://huggingface.co/datasets/DebdipCS/Latent-Resonance-AI-Image-Forensics-Benchmark-N100.clker-images
Dataset Card for Clker.com Images
Dataset Summary
This dataset contains 140,313 public domain clipart images collected from Clker.com. Clker.com hosts user-shared vector clip art that is explicitly released into the public domain (CC0). The dataset includes the images themselves along with metadata such as titles and tags associated with each image.
Languages
The dataset is primarily monolingual:
English (en): All image titles and tags are in… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/clker-images.rule34lol-images-part1
Dataset Card for rule34lol-images-part1
Dataset Summary
This dataset contains information about image files from rule34.lol, a booru-style imageboard. The dataset includes metadata for 196,000 image files, including URLs, tags, file information, and like counts. The actual image files are stored in zip archives, with each archive containing 1000 image files. This is Part 1 of 2 for the complete rule34lol-images dataset. Part 2 can be found here.
Languages… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/rule34lol-images-part1.AVAINT-IMGOpenHotels-Updated
OpenHotels-Updated
OpenHotels-Updated is a large-scale hotel image retrieval benchmark built from hotel-room imagery and associated hotel metadata. The dataset is designed for hotel-scale retrieval: given a query image, a system must retrieve the matching hotel from a large gallery containing both true matching classes and many distractor hotel classes.
Dataset Structure
The release contains tar-sharded image files under shards/ and four metadata files:
shards/… See the full description on the dataset page: https://huggingface.co/datasets/imagingforgood/OpenHotels-Updated.ikea-us-products-2025
IKEA US Product Dataset (July 2025)
This dataset is a structured snapshot of ~30,000 IKEA US products, scraped from the official IKEA US website in July 2025.
It contains product metadata (titles, descriptions, categories, materials, care instructions, etc.) and associated product images.
Contents
products-us.jsonl — one JSON object per product with structured fields.
images-us/ — the first "hero" image for each product, downloaded via image_downloader_first.py.… See the full description on the dataset page: https://huggingface.co/datasets/doniariz/ikea-us-products-2025.aesthetic-corpus
InspiredHub Aesthetic Corpus — Open Layer
A structured dataset of 81,911 artworks, 11,568 books, 220 composers, and 146 philosophy pages from the world's major museums and archives.
Overview
Collection
Items
Description
Artworks
81,911
Paintings, calligraphy, sculpture, ceramics from 15+ museums
Books
11,568
Classic literature, philosophy, poetry (840 philosophy texts)
Composers
220
Classical composers with works catalog
Philosophy Pages
146… See the full description on the dataset page: https://huggingface.co/datasets/InspiredHub/aesthetic-corpus.solar-flare-hmi-datasplitsThis dataset is intended to be used for training/testing solar flare forecasting models. It contains various data splits (in json format) of SDO/HMI magnetogram images compiled by
Boucheron, L.E., et al., 2023, Sci Data 10, 825, https://doi.org/10.1038/s41597-023-02628-8.
Splits "train", "val", "test" corresponds to the original data splits provided by Boucheron et al., while the other splits are created by downsampling the No-Flare and C flare class to obtain
more balanced splits and… See the full description on the dataset page: https://huggingface.co/datasets/inaf-oact-ai/solar-flare-hmi-datasplits.OpenHotels
OpenHotels
OpenHotels is a large-scale hotel image retrieval benchmark built from hotel-room imagery and associated hotel metadata. The dataset is designed for hotel-scale retrieval: given a query image, a system must retrieve the matching hotel from a large gallery containing both true matching classes and many distractor hotel classes.
Dataset Structure
The release contains tar-sharded image files under shards/ and four metadata files:
shards/… See the full description on the dataset page: https://huggingface.co/datasets/imagingforgood/OpenHotels.sigui-depin-1m
Sigui DePIN 1M — Multichain Transaction Graph Dataset
The largest open dataset of annotated blockchain transaction graph visualizations for AI security research.
📋 Dataset Description
sigui-depin-1m contains 1,000,000 visual transaction graph images generated from 1.87 million real on-chain transactions across Ethereum, Arbitrum, and Polygon. Each image is annotated with an attack topology label designed to train Vision-Language Models to detect DeFi threats.… See the full description on the dataset page: https://huggingface.co/datasets/Ibonon/sigui-depin-1m.Agricultural-Pest-Disease-Image-Annotation-Bounding-Box-Polygon-Detection
Agricultural Pest & Disease Image Annotation (Bounding Box/Polygon Detection)
For intelligent pest and disease recognition in the agriculture and livestock industry, this dataset provides target region annotations on raw plant/disease images. Annotation results support both bounding boxes and polygon shapes, enabling models to learn object locations and fine-grained contours. It covers common pest/disease targets and supplies strong supervision for detection and segmentation… See the full description on the dataset page: https://huggingface.co/datasets/shangzx/Agricultural-Pest-Disease-Image-Annotation-Bounding-Box-Polygon-Detection.Asparagus-Identification-Dataset
Asparagus Identification Dataset
The current agricultural industry faces challenges in efficient crop monitoring and recognition. Traditional manual detection methods are inefficient and prone to errors. Existing solutions often rely on empirical judgment without scientific model support. This dataset aims to provide diverse asparagus images to help train automatic recognition models, improving the accuracy and efficiency of crop monitoring. The dataset includes images of asparagus… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Asparagus-Identification-Dataset.Experiment-Design-Sketch-Image-Classification-Dataset
Experiment Design Sketch Image Classification Dataset
In the field of industrial manufacturing, the design process often relies on a large number of design sketches that need to be quickly converted into actual engineering designs during the subsequent manufacturing stages. However, manually processing these sketches is often time-consuming and prone to errors, currently relying mainly on manual labeling and conversion by designers, which is inefficient and unstable. Existing… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Experiment-Design-Sketch-Image-Classification-Dataset.FTAG
FTAG: Fabric aTtribute Analysis of Garments
FTAG is a dataset of 12,304 women's dresses (ASOS product photos) annotated with
three fabric attributes derived from the retailer's product composition data:
Material composition — e.g. Main: 95% Polyester, 5% Elastane (31 canonical materials)
Structure type — woven / knit / others (3 classes)
Fabric family — e.g. jersey, chiffon, denim (15 general families)
The task: predict all three attributes from a single garment image.
Images… See the full description on the dataset page: https://huggingface.co/datasets/image2garment/FTAG.generated-imagenette
Generated Imagenette Dataset
Description
This repository contains the dataset used for the generative-data-augmentation project. The dataset is organized as follows:
Dataset Structure
analysis/: This directory contains analysis related to the dataset.
metadata/: This directory contains the list of file path used for the Synthetic (Noisy) and Synthetic (Clean) datasets.
synthetic/: This directory contains the image files. Each folder represents a class.… See the full description on the dataset page: https://huggingface.co/datasets/czl/generated-imagenette.generated-imagewoof
Generated Imagewoof Dataset
Description
This repository contains the dataset used for the generative-data-augmentation project. The dataset is organized as follows:
Dataset Structure
analysis/: This directory contains analysis related to the dataset.
metadata/: This directory contains the list of file path used for the Synthetic (Noisy) and Synthetic (Clean) datasets.
synthetic/: This directory contains the image files. Each folder represents a class.… See the full description on the dataset page: https://huggingface.co/datasets/czl/generated-imagewoof.
