datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gpt-image-2-prompts-datasets
🖼️ GPT Image 2 Prompt Dataset
🖼️ The ultimate GPT Image 2 prompt dataset (5GB+). 15,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators.
This project is a massive collection of prompts used for OpenAI's GPT Image 2 model and the resulting generated images. The entire dataset exceeds 5GB and contains 15,000+ images, all structured into a comprehensive dataset.
Due to… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/gpt-image-2-prompts-datasets.dummy_image_text_data
Dataset Card for "dummy_image_text_data"
More Information needed
image-segmentation-toy-dataopen-image-preferences-v1
Open Image Preferences
Prompt: Anime-style concept art of a Mayan Quetzalcoatl biomutant, dystopian world, vibrant colors, 4K.
Image 1
Image 2
Prompt: 8-bit pixel art of a blue knight, green car, and glacier landscape in Norway, fantasy style, colorful and detailed.
Image 1… See the full description on the dataset page: https://huggingface.co/datasets/data-is-better-together/open-image-preferences-v1.SukaSuka-image-dataset
该数据集包含了《末日时在做什么?有没有空?可以来拯救吗?》大部分主要角色角色的图像数据,来源为动漫截图与同人二创。
为方便LoRA模型训练,所有图片尺寸均截为512x640尺寸,相应打标主要由Waifu Diffusion 1.4 Tagger V2自动完成,部分手工调整。
欢迎提交PR补充或修正本数据集!
Alpha:8.21号之后的clone都是放大了两倍的图片,这是为了sdxl做准备,如果你还需要512*640尺寸的数据集,你可以在clone之后,执行下面的命令
git checkout 183e253c4c304fc6c5ef5046f1940712c349c94e
相关数据集的更正作业正在火热的进行中,请期待继续的更新吧~
Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-ForecastingThe sp500stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 4,213 S&P 500 stocks.
The hs300stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 858 HS 300 stocks.
If you find our research helpful, please cite our paper:
@article{xu2025finmultitime,
title={FinMultiTime: A Four-Modal Bilingual Dataset for… See the full description on the dataset page: https://huggingface.co/datasets/Wenyan0110/Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-Forecasting.Defactify_Image_Dataset
Defactify_Image_Dataset
This dataset is associated with the paper A Comprehensive Dataset for Human vs. AI Generated Image Detection.
📝 Dataset Description
Dataset Summary
The Defactify_Image_Dataset (A Comprehensive Dataset for Human vs. AI Generated Image Detection) is a high-quality collection of 96,000 images and associated metadata designed to benchmark models for detecting and identifying the source of artificially generated content. Built using the MS… See the full description on the dataset page: https://huggingface.co/datasets/Rajarshi-Roy-research/Defactify_Image_Dataset.Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-DatasetTunisian Proverbs with Image Associations: A Cultural and Linguistic Dataset
Description
This dataset explores the rich oral tradition of Tunisian proverbs mapped into text format, pairing each with contextual explanations, English translations both word-to-word and it's equivalent Target Language dynamic, Automated prompt and AI-generated visual interpretations.
It bridges linguistic, cultural, and visual modalities making it valuable for tasks in cross-cultural NLP, generative… See the full description on the dataset page: https://huggingface.co/datasets/HabibaAbderrahim/Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-Dataset.MLLM-Generated-Image-Detection-Dataset
MLLM-Generated Image Dataset
This dataset contains real and AI-generated image samples organized for binary MLLM-generated image detection.
Paper | Code
Dataset Summary
We construct an MLLM-generated image detection benchmark from GPT Image2 and Nano Banana2. This benchmark covers texture-dominated, structure-dominated, and hybrid-dominated. It is designed to evaluate detector performance under the new challenges introduced by large-scale image generation models.… See the full description on the dataset page: https://huggingface.co/datasets/zr-zhang/MLLM-Generated-Image-Detection-Dataset.Crop_Disease_Image_Dataset
Crop Disease Image Dataset (5 Crops, 19 Classes)
Dataset Summary
The Crop Disease Image Dataset is a curated, high-quality agricultural image dataset designed for computer vision, deep learning, and smart farming applications. It contains 22,169 RGB leaf images spanning 5 major crops across 19 distinct healthy and diseased classes.
This dataset was constructed by collecting, filtering, and standardizing images from multiple open-source agricultural repositories… See the full description on the dataset page: https://huggingface.co/datasets/ipartzix/Crop_Disease_Image_Dataset.malfunction-image-datasetdummy_image_class_data
Dataset Card for "dummy_image_class_data"
More Information needed
T2I-RiskyPrompt-ImageDataset
T2I-RiskyPrompt-Derived Images
Overview
This image dataset is derived from the project T2I-RiskyPrompt. T2I-RiskyPrompt provides a hierarchical risk taxonomy (6 primary categories and 14 subcategories) and a set of 6,432 human-validated risky prompts, where each prompt is annotated with hierarchical labels and detailed risk reasons. This image dataset contains 20,373 images generated from SD3 and FLUX using T2I-RiskyPrompt, where each image is annotated with both… See the full description on the dataset page: https://huggingface.co/datasets/datarr/T2I-RiskyPrompt-ImageDataset.Dataset-Image-1-Heir
📦 Heir-Image - Changelog
Date
Dataset
Details
08/28
Dataset-Image
1997 images - initial release
08/31
Dataset-Image2
873 images - updated version, higher quality than Dataset-Image :))
08/31
Dataset-Image2.1
610 high-quality images "Lmao"
image-relighting-diffusion-data
Learning Illumination Control in Diffusion Models — Dataset (HF)
Public data and evaluation assets for Learning Illumination Control in Diffusion Models (ReALM-GEN @ ICLR 2026).
Code
github.com/nishitanand/image-relighting-diffusion
Model weights
huggingface.co/nishitanand/sd-image-relighting-model
Paper
arxiv.org/abs/2604.24877
Project site
nishitanand.github.io/relighting-diffusion-website
Download (CLI)
Install the Hugging Face CLI… See the full description on the dataset page: https://huggingface.co/datasets/nishitanand/image-relighting-diffusion-data.image-dataCIFAKE-image-datasetai-image-detector-dataset
AI Image Detector Dataset (training-v1)
wkaandemir/ai-image-detector modelini eğitmek için kullanılan, 20.000 normalize edilmiş görselden oluşan dengeli ve kaynak-farkında (source-aware) bir görüntü sınıflandırma veri kümesi. Görev, görselleri gerçek (real) ve yapay (fake) olarak ikili sınıflandırmaktır.
Bu veri kümesi, modelin kalibrasyon ve eşik seçiminde kullanılmayan, genelleme ölçümü için kaynak bazlı ayrı tutulmuş bir external_test split'i de içerir.
📌… See the full description on the dataset page: https://huggingface.co/datasets/wkaandemir/ai-image-detector-dataset.Chinese_Children_Image_Captioning_Dataset_Split0
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split0.fictional-characters-image-datasetHow to use
Here is how to use this dataset:
from datasets import load_dataset
dataset = load_dataset("gryffindor-ISWS/fictional-characters-image-dataset")
This repository contains fictional characters dataset constructed from Wikidata for the research project "Draw Me Like Your Triples: Leveraging Generative AI for the Completion of Wikidata". The project was conducted by Raia Abu Ahmad, Martin Critelli, Şefika Efeoğlu, Eleonora Mancini, Célian Ringwald and Xinyue Zhang under the… See the full description on the dataset page: https://huggingface.co/datasets/gryffindor-ISWS/fictional-characters-image-dataset.cat-image-datasetLabels (in this order):
sks cat sitting on a chair in front of a box of chocolates
sks cat playing on the Steam Deck
sks cat wearing a necklace while sitting in a box on a sofa
a box with four donuts in front of sks cat
sks cat wearing a pink veil with flowers on it and a dagger made out of yellow cardboard
sks cat between two pillows, with one pillow showing a polar bear and the other a fox
sks cat with an espresso reading the newspaper
a close up of a hand petting sks cat on the head
sks cat… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/cat-image-dataset.face-segmentation-image-dataset
Image Dataset of Face Segmentation for recognition tasks
Dataset comprises 87,800+ images annotated with 100+ landmarks, providing a comprehensive foundation for research in face recognition, segmentation tasks, and object recognition. It is designed to support the development of learning models, recognition algorithms, and segmentation techniques.
By utilizing this dataset, researchers and developers can advance their understanding and capabilities in facial recognition, face… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/face-segmentation-image-dataset.pagoda-text-and-image-dataset
Dataset Card for "pagoda-text-and-image-dataset"
More Information needed
yt_full_image_dataset
Dataset Card for "yt_full_image_dataset"
More Information needed
drill-core-image-dataset
Dataset Card for Drill Core Image Dataset (DCID)
Dataset Details
Dataset Description
The Drill Core Image Dataset (DCID) is a large-scale benchmark designed for lithology classification based on RGB core images. It provides two primary versions:
DCID-7: 7 lithology categories with 5,000 images per class.
DCID-35: 35 lithology categories with 1,000 images per class.
All original images are 512×512 pixels in resolution. Each category is split into… See the full description on the dataset page: https://huggingface.co/datasets/168sir/drill-core-image-dataset.image-matching-test-datasetpublic_image_dataSterling_AI_Image_Dataset
Sterling Ai Image Dataset
This is a compiled categorized dataset for anyone but to be used in my ai project.
What it is?
This is a dataset were I took alot of online images and put them in folders that discribe the object the photo is centered on
to be used in whatever u can find to use it in.
Image Takedown Request
In the works
Category Recomendations
I have limited ideas on what to put in here i will find a way but feel free to… See the full description on the dataset page: https://huggingface.co/datasets/Sterling-Ai/Sterling_AI_Image_Dataset.fashion-image-datasetChinese_Children_Image_Captioning_Dataset_Split1
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split1.
