CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Goku-OpenLab /gpt-image-2-prompts-datasets 🖼️ GPT Image 2 Prompt Dataset 🖼️ The ultimate GPT Image 2 prompt dataset (5GB+). 15,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators. This project is a massive collection of prompts used for OpenAI's GPT Image 2 model and the resulting generated images. The entire dataset exceeds 5GB and contains 15,000+ images, all structured into a comprehensive dataset. Due to… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/gpt-image-2-prompts-datasets.imagetext-to-image10K<n<100K6 likes78k downloads1mo agoHugging Face02hf-internal-testing /dummy_image_text_data Dataset Card for "dummy_image_text_data" More Information needed imagen<1K1 likes11k downloads4y agoHugging Face03nielsr /image-segmentation-toy-dataimagen<1K0 likes9.4k downloads4y agoHugging Face04data-is-better-together /open-image-preferences-v1 Open Image Preferences Prompt: Anime-style concept art of a Mayan Quetzalcoatl biomutant, dystopian world, vibrant colors, 4K. Image 1 Image 2 Prompt: 8-bit pixel art of a blue knight, green car, and glacier landscape in Norway, fantasy style, colorful and detailed. Image 1… See the full description on the dataset page: https://huggingface.co/datasets/data-is-better-together/open-image-preferences-v1.imagetext-to-image1K<n<10K31 likes4.9k downloads2y agoHugging Face05Carzit /SukaSuka-image-dataset 该数据集包含了《末日时在做什么?有没有空?可以来拯救吗?》大部分主要角色角色的图像数据,来源为动漫截图与同人二创。 为方便LoRA模型训练,所有图片尺寸均截为512x640尺寸,相应打标主要由Waifu Diffusion 1.4 Tagger V2自动完成,部分手工调整。 欢迎提交PR补充或修正本数据集! Alpha:8.21号之后的clone都是放大了两倍的图片,这是为了sdxl做准备,如果你还需要512*640尺寸的数据集,你可以在clone之后,执行下面的命令 git checkout 183e253c4c304fc6c5ef5046f1940712c349c94e 相关数据集的更正作业正在火热的进行中,请期待继续的更新吧~ imagen<1K5 likes3k downloads1y agoHugging Face06Wenyan0110 /Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-ForecastingThe sp500stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 4,213 S&P 500 stocks. The hs300stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 858 HS 300 stocks. If you find our research helpful, please cite our paper: @article{xu2025finmultitime, title={FinMultiTime: A Four-Modal Bilingual Dataset for… See the full description on the dataset page: https://huggingface.co/datasets/Wenyan0110/Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-Forecasting.imagen<1K12 likes2.7k downloads1y agoHugging Face07Rajarshi-Roy-research /Defactify_Image_Dataset Defactify_Image_Dataset This dataset is associated with the paper A Comprehensive Dataset for Human vs. AI Generated Image Detection. 📝 Dataset Description Dataset Summary The Defactify_Image_Dataset (A Comprehensive Dataset for Human vs. AI Generated Image Detection) is a high-quality collection of 96,000 images and associated metadata designed to benchmark models for detecting and identifying the source of artificially generated content. Built using the MS… See the full description on the dataset page: https://huggingface.co/datasets/Rajarshi-Roy-research/Defactify_Image_Dataset.imageimage-classification10K<n<100K22 likes2.4k downloads4mo agoHugging Face08HabibaAbderrahim /Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-DatasetTunisian Proverbs with Image Associations: A Cultural and Linguistic Dataset Description This dataset explores the rich oral tradition of Tunisian proverbs mapped into text format, pairing each with contextual explanations, English translations both word-to-word and it's equivalent Target Language dynamic, Automated prompt and AI-generated visual interpretations. It bridges linguistic, cultural, and visual modalities making it valuable for tasks in cross-cultural NLP, generative… See the full description on the dataset page: https://huggingface.co/datasets/HabibaAbderrahim/Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-Dataset.imagetranslationn<1K0 likes2.2k downloads1y agoHugging Face09zr-zhang /MLLM-Generated-Image-Detection-Dataset MLLM-Generated Image Dataset This dataset contains real and AI-generated image samples organized for binary MLLM-generated image detection. Paper | Code Dataset Summary We construct an MLLM-generated image detection benchmark from GPT Image2 and Nano Banana2. This benchmark covers texture-dominated, structure-dominated, and hybrid-dominated. It is designed to evaluate detector performance under the new challenges introduced by large-scale image generation models.… See the full description on the dataset page: https://huggingface.co/datasets/zr-zhang/MLLM-Generated-Image-Detection-Dataset.imageimage-classification1K<n<10K1 likes2k downloads2mo agoHugging Face10ipartzix /Crop_Disease_Image_Dataset Crop Disease Image Dataset (5 Crops, 19 Classes) Dataset Summary The Crop Disease Image Dataset is a curated, high-quality agricultural image dataset designed for computer vision, deep learning, and smart farming applications. It contains 22,169 RGB leaf images spanning 5 major crops across 19 distinct healthy and diseased classes. This dataset was constructed by collecting, filtering, and standardizing images from multiple open-source agricultural repositories… See the full description on the dataset page: https://huggingface.co/datasets/ipartzix/Crop_Disease_Image_Dataset.imageimage-classification1K<n<10K1 likes1.6k downloads2mo agoHugging Face11irene93 /malfunction-image-datasetimage10K<n<100K0 likes1.5k downloads1y agoHugging Face12hf-internal-testing /dummy_image_class_data Dataset Card for "dummy_image_class_data" More Information needed imagen<1K0 likes1.2k downloads4y agoHugging Face13datarr /T2I-RiskyPrompt-ImageDataset T2I-RiskyPrompt-Derived Images Overview This image dataset is derived from the project T2I-RiskyPrompt. T2I-RiskyPrompt provides a hierarchical risk taxonomy (6 primary categories and 14 subcategories) and a set of 6,432 human-validated risky prompts, where each prompt is annotated with hierarchical labels and detailed risk reasons. This image dataset contains 20,373 images generated from SD3 and FLUX using T2I-RiskyPrompt, where each image is annotated with both… See the full description on the dataset page: https://huggingface.co/datasets/datarr/T2I-RiskyPrompt-ImageDataset.imagetext-to-image2 likes1.1k downloads10mo agoHugging Face14RoxERL0002113 /Dataset-Image-1-Heir 📦 Heir-Image - Changelog Date Dataset Details 08/28 Dataset-Image 1997 images - initial release 08/31 Dataset-Image2 873 images - updated version, higher quality than Dataset-Image :)) 08/31 Dataset-Image2.1 610 high-quality images "Lmao" imageimage-classification1K<n<10K1 likes1k downloads26d agoHugging Face15nishitanand /image-relighting-diffusion-data Learning Illumination Control in Diffusion Models — Dataset (HF) Public data and evaluation assets for Learning Illumination Control in Diffusion Models (ReALM-GEN @ ICLR 2026). Code github.com/nishitanand/image-relighting-diffusion Model weights huggingface.co/nishitanand/sd-image-relighting-model Paper arxiv.org/abs/2604.24877 Project site nishitanand.github.io/relighting-diffusion-website Download (CLI) Install the Hugging Face CLI… See the full description on the dataset page: https://huggingface.co/datasets/nishitanand/image-relighting-diffusion-data.imageimage-to-image10K<n<100K0 likes931 downloads3mo agoHugging Face16junaid17 /image-dataimage1K<n<10K0 likes894 downloads9mo agoHugging Face17dragonintelligence /CIFAKE-image-datasetimage100K<n<1M0 likes884 downloads2y agoHugging Face18wkaandemir /ai-image-detector-dataset AI Image Detector Dataset (training-v1) wkaandemir/ai-image-detector modelini eğitmek için kullanılan, 20.000 normalize edilmiş görselden oluşan dengeli ve kaynak-farkında (source-aware) bir görüntü sınıflandırma veri kümesi. Görev, görselleri gerçek (real) ve yapay (fake) olarak ikili sınıflandırmaktır. Bu veri kümesi, modelin kalibrasyon ve eşik seçiminde kullanılmayan, genelleme ölçümü için kaynak bazlı ayrı tutulmuş bir external_test split'i de içerir. 📌… See the full description on the dataset page: https://huggingface.co/datasets/wkaandemir/ai-image-detector-dataset.imageimage-classification10K<n<100K0 likes882 downloads2mo agoHugging Face19svjack /Chinese_Children_Image_Captioning_Dataset_Split0 CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions) CODP-1200: An AIGC based benchmark for assisting in child language acquisition 数据集介绍 目前已知最大的儿童图像描述数据集,children image captioning 共有1200张图片 每张图片对应五个中文描述,每两张图片为一组 描述文字600*5=3000 如果使用CODP-1200数据集,请引用以下文章 @article{LENG2024102627, title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition}, journal = {Displays}, volume = {82}, pages = {102627}, year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split0.image1K<n<10K0 likes861 downloads2y agoHugging Face20gryffindor-ISWS /fictional-characters-image-datasetHow to use Here is how to use this dataset: from datasets import load_dataset dataset = load_dataset("gryffindor-ISWS/fictional-characters-image-dataset") This repository contains fictional characters dataset constructed from Wikidata for the research project "Draw Me Like Your Triples: Leveraging Generative AI for the Completion of Wikidata". The project was conducted by Raia Abu Ahmad, Martin Critelli, Şefika Efeoğlu, Eleonora Mancini, Célian Ringwald and Xinyue Zhang under the… See the full description on the dataset page: https://huggingface.co/datasets/gryffindor-ISWS/fictional-characters-image-dataset.imagen<1K0 likes736 downloads3y agoHugging Face21peft-internal-testing /cat-image-datasetLabels (in this order): sks cat sitting on a chair in front of a box of chocolates sks cat playing on the Steam Deck sks cat wearing a necklace while sitting in a box on a sofa a box with four donuts in front of sks cat sks cat wearing a pink veil with flowers on it and a dagger made out of yellow cardboard sks cat between two pillows, with one pillow showing a polar bear and the other a fox sks cat with an espresso reading the newspaper a close up of a hand petting sks cat on the head sks cat… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/cat-image-dataset.imagen<1K0 likes627 downloads4mo agoHugging Face22UniDataPro /face-segmentation-image-dataset Image Dataset of Face Segmentation for recognition tasks Dataset comprises 87,800+ images annotated with 100+ landmarks, providing a comprehensive foundation for research in face recognition, segmentation tasks, and object recognition. It is designed to support the development of learning models, recognition algorithms, and segmentation techniques. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in facial recognition, face… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/face-segmentation-image-dataset.imagevideo-classificationn<1K5 likes524 downloads1mo agoHugging Face23nojiyoon /pagoda-text-and-image-dataset Dataset Card for "pagoda-text-and-image-dataset" More Information needed imagen<1K1 likes516 downloads3y agoHugging Face24vargr /yt_full_image_dataset Dataset Card for "yt_full_image_dataset" More Information needed image100K<n<1M1 likes509 downloads3y agoHugging Face25168sir /drill-core-image-dataset Dataset Card for Drill Core Image Dataset (DCID) Dataset Details Dataset Description The Drill Core Image Dataset (DCID) is a large-scale benchmark designed for lithology classification based on RGB core images. It provides two primary versions: DCID-7: 7 lithology categories with 5,000 images per class. DCID-35: 35 lithology categories with 1,000 images per class. All original images are 512×512 pixels in resolution. Each category is split into… See the full description on the dataset page: https://huggingface.co/datasets/168sir/drill-core-image-dataset.imageimage-classification10K<n<100K3 likes500 downloads1y agoHugging Face26hf-internal-testing /image-matching-test-datasetimagen<1K0 likes447 downloads2y agoHugging Face27laiviet /public_image_dataimage0 likes411 downloads11mo agoHugging Face28Sterling-Ai /Sterling_AI_Image_Dataset Sterling Ai Image Dataset This is a compiled categorized dataset for anyone but to be used in my ai project. What it is? This is a dataset were I took alot of online images and put them in folders that discribe the object the photo is centered on to be used in whatever u can find to use it in. Image Takedown Request In the works Category Recomendations I have limited ideas on what to put in here i will find a way but feel free to… See the full description on the dataset page: https://huggingface.co/datasets/Sterling-Ai/Sterling_AI_Image_Dataset.imagen<1K1 likes382 downloads2mo agoHugging Face29Aayush672 /fashion-image-datasetimage10K<n<100K0 likes377 downloads1y agoHugging Face30svjack /Chinese_Children_Image_Captioning_Dataset_Split1 CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions) CODP-1200: An AIGC based benchmark for assisting in child language acquisition 数据集介绍 目前已知最大的儿童图像描述数据集,children image captioning 共有1200张图片 每张图片对应五个中文描述,每两张图片为一组 描述文字600*5=3000 如果使用CODP-1200数据集,请引用以下文章 @article{LENG2024102627, title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition}, journal = {Displays}, volume = {82}, pages = {102627}, year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split1.image1K<n<10K0 likes367 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.