CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nyuuzyou /OpenGameArt-CC0 Dataset Card for OpenGameArt-CC0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons 0 (CC0) license, making them effectively public domain works. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual:… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC0.audioimage-classification10K<n<100K10 likes1.6k downloads1y agoHugging Face02QCRI /CrisisMMD CrisisMMD: Multimodal Twitter Datasets from Natural Disasters The CrisisMMD multimodal Twitter dataset consists of several thousand manually annotated tweets and images collected during seven major natural disasters, including earthquakes, hurricanes, wildfires, and floods from 2017. The dataset includes three types of annotations: On HuggingFace, we hosted version 2.0 of the CrisisMMD dataset. Please see further information below. Disaster Response Tasks Task 1:… See the full description on the dataset page: https://huggingface.co/datasets/QCRI/CrisisMMD.imageimage-classification10K<n<100K5 likes1.5k downloads2y agoHugging Face03neurlang /Minecraft-Skins-Captioned-1M Dataset Card for Minecraft Skins Dataset Summary This dataset contains 981,079 unique Minecraft player skins collected from various sources. Each skin is stored as a base64-encoded image with a unique identifier. Dataset Structure Data Fields This dataset includes the following fields: hash: A data dependent hash. These hashes are generated from raw bytes and will be same if the skin is identical. image: The skin image encoded in base64 format.… See the full description on the dataset page: https://huggingface.co/datasets/neurlang/Minecraft-Skins-Captioned-1M.textimage-classification1M<n<10M7 likes689 downloads11mo agoHugging Face04nyuuzyou /OpenGameArt-CC-BY-SA-3.0 Dataset Card for OpenGameArt-CC-BY-SA-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution-ShareAlike 3.0 Unported (CC-BY-SA-3.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-SA-3.0.audioimage-classification1K<n<10K1 likes582 downloads1y agoHugging Face05DSULT-Chiharu /the-un-laion-templeAll files uploaded. Enjoy! Dataset Card for The Unlaion Temple Dataset Details Dataset Description Laion-5B is still not public, so we decided to create our own dataset. The Unlaion Temple is a raw dataset of CommonCrawl images (Estimated to be a total of 2 Billion urls). We haven't verified whether the links in this dataset are functional. You are responsible for handling the data. We've made some improvements to the dataset based on user feedback: All… See the full description on the dataset page: https://huggingface.co/datasets/DSULT-Chiharu/the-un-laion-temple.textimage-classification1B<n<10B2 likes544 downloads2y agoHugging Face06IAmFuch /viet-cultural-vqa 🇻🇳 Vietnamese Cultural VQA Dataset 📖 Dataset Description The Vietnamese Cultural VQA Dataset is a comprehensive multimodal dataset designed for Visual Question Answering (VQA) tasks focused on Vietnamese cultural heritage. This dataset aims to bridge the gap in understanding and preserving Vietnamese culture through AI-powered visual understanding and question answering. 🎯 Dataset Summary 📊 Total Images: 28,505 high-quality cultural images 💬 Total… See the full description on the dataset page: https://huggingface.co/datasets/IAmFuch/viet-cultural-vqa.imagevisual-question-answering10K<n<100K0 likes539 downloads5mo agoHugging Face07nyuuzyou /OpenGameArt-CC-BY-3.0 Dataset Card for OpenGameArt-CC-BY-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 3.0 (CC-BY-3.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual: English (en): All… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-3.0.textimage-classification1K<n<10K1 likes395 downloads1y agoHugging Face08yuzhench /glaucoma-expert-cot-raw-1077 Glaucoma Expert Chain-of-Thought Ophthalmologist six-step reasoning reports for fundus photographs, each paired with a binary glaucoma label. 1,074 cases from LAG and Papila. Files file rows split expert_cot_trainval.jsonl 915 train (823) + val (92) expert_cot_test.jsonl 159 test images/ 1,074 <source>_<id>.jpg Record schema { "id": "1689", "source": "LAG", "image": "LAG_1689.jpg", "split": "train"… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/glaucoma-expert-cot-raw-1077.imageimage-classificationn<1K0 likes314 downloads2mo agoHugging Face09nyuuzyou /clker-svg Dataset Card for Clker.com SVG Images Dataset Summary This dataset contains 255,758 public domain SVG vector clipart images collected from Clker.com. Clker.com hosts user-shared vector clip art that is explicitly released into the public domain (CC0). The dataset includes the SVG content itself along with metadata such as titles and tags associated with each image. The SVG files in this dataset have been minified using tdewolff/minify to reduce file size while… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/clker-svg.imageimage-classification100K<n<1M5 likes272 downloads1y agoHugging Face10opendiffusionai /cc12m-cleaned CC12m-cleaned This dataset builds on two others: The Conceptual Captions 12million dataset, which lead to the LLaVa captioned subset done by CaptionEmporium (The latter is the same set, but swaps out the (Conceptual Captions 12million) often-useless alt-text captioning for decent ones_ I have then used the llava captions as a base, and used the detailed descrptions to filter out images with things like watermarks, artist signatures, etc. I have also manually thrown out all… See the full description on the dataset page: https://huggingface.co/datasets/opendiffusionai/cc12m-cleaned.imagetext-to-image1M<n<10M13 likes269 downloads2y agoHugging Face11eelianafang /MOUNT-Cattle Updates/News 📣 🎉 News (Feb. 2026): The dataset paper FSMC-Pose has been accepted for CVPR 2026 Findings! 🔗 News: Please find the open-source dataset on Hugging Face: MOUNT-Cattle. 🔥 Downloads reached 2.4k within 7 days of release.       📌 Overview Mounting posture is an important visual indicator of estrus in dairy cattle. MOUNT-Cattle is a mounting dataset, covering 1,176 mounting instances, which follows the COCO format… See the full description on the dataset page: https://huggingface.co/datasets/eelianafang/MOUNT-Cattle.imageobject-detection1K<n<10K5 likes268 downloads6mo agoHugging Face12DesanSilva /sscc-compact-av SSCC compact balanced multimodal subset This private derived dataset contains 288 synchronized SSCC clips from 15 medium-load, clean operating conditions at speeds 60, 80, and 100. It retains recorder FLAC audio, four anti-aliased 25 kHz vibration channels in compressed float32 NPZ, and five sparse frames from both iOS and Android videos. Five sample IDs retain both unchanged source MP4s for presentation and loader tests. The subset is balanced between normal and fault states… See the full description on the dataset page: https://huggingface.co/datasets/DesanSilva/sscc-compact-av.textimage-classificationn<1K0 likes197 downloads14d agoHugging Face13caiotheodoro /vernier vernier Error bars on a dataset vendor's quality claim. Build AI publishes hand-visibility and active-manipulation rates for Egocentric-10K / Egocentric-100K, judged once by gemini-2.5-flash with no human gold, no interval, and no test that the judge scores a factory floor and a home kitchen on the same scale. This release is the data behind an independent, pre-registered measurement of that claim: human labels against a written rubric, a live open-weights judge on the same… See the full description on the dataset page: https://huggingface.co/datasets/caiotheodoro/vernier.tabularimage-classification10K<n<100K0 likes180 downloads18d agoHugging Face14chen-yingfa /CHUBS CHUBS: A Large-Scale Dataset of Chu Bamboo Slip Script Code | Paper (upcoming) Introduction This is a large-scale dataset of Chu bamboo slip (CBS, Chinese: 楚简, chujian) script, an ancient Chinese script used during the Spring and Autumn period over 2,000 years ago. This dataset consists of two parts: The main dataset where each example is an image and the corresponding text label. This part is contained in the glyphs.zip ZIP file. A character detection dataset… See the full description on the dataset page: https://huggingface.co/datasets/chen-yingfa/CHUBS.imageimage-classification1K<n<10K1 likes175 downloads2y agoHugging Face15auniquesun /Point-CacheThe datasets in this repository are used in the paper Point-Cache: Test-time Dynamic and Hierarchical Cache for Robust and Generalizable Point Cloud Analysis. Datasets The folder structure of used datasets should be organized as follows. /path/to/Point-Cache |----data # placed in the same level as `runners`, `scripts`, etc. |----modelnet_c |----sonn_c |----obj_bg |----obj_only |----hardest |----modelnet40… See the full description on the dataset page: https://huggingface.co/datasets/auniquesun/Point-Cache.textimage-classification1K<n<10K1 likes164 downloads1y agoHugging Face16annoymous-1 /CC-Bench CC-Bench: A Cognitive Conflict Benchmark for MLLMs in Safety-Critical Visual Inspection CC-Bench is a joint medical-industrial benchmark for evaluating whether multimodal large language models (MLLMs) remain visually grounded when plausible textual context conflicts with image evidence. The benchmark reorganizes public anomaly datasets into a unified four-way multiple-choice QA format for high-risk visual inspection. This repository currently contains: 4,282 images in total 2,157… See the full description on the dataset page: https://huggingface.co/datasets/annoymous-1/CC-Bench.imagevisual-question-answering1K<n<10K0 likes130 downloads5mo agoHugging Face17nyuuzyou /OpenGameArt-CC-BY-4.0 Dataset Card for OpenGameArt-CC-BY-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 4.0 International (CC-BY-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual: English… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-4.0.audioimage-classification1K<n<10K2 likes123 downloads1y agoHugging Face18harvardairobotics /Fundus-CoT Glaucoma Expert Chain-of-Thought Ophthalmologist six-step reasoning reports for fundus photographs, each paired with a binary glaucoma label. 1,074 cases from LAG and Papila. Files file rows glaucoma / not train.jsonl 823 304 / 519 val.jsonl 92 46 / 46 test.jsonl 159 79 / 80 images/ 1,074 <source>_<id>.jpg Record schema { "id": "1689", "source": "LAG", "image": "LAG_1689.jpg", "split": "train", "final_diagnosis_GT":… See the full description on the dataset page: https://huggingface.co/datasets/harvardairobotics/Fundus-CoT.imageimage-classification1K<n<10K0 likes115 downloads2mo agoHugging Face19VatsaDev /cnn_muffins CNN Muffins A compact dog-versus-muffin image-classification dataset built around the well-known visual confusion between Chihuahua faces and blueberry muffins. Dataset structure Split Dogs Muffins Total Train 319 161 480 Validation 36 18 54 Hard-16 benchmark 8 8 16 The hard-16 benchmark is isolated from train and validation. The JSONL files use repository-relative image paths: The benchmark labels follow the original 4x4 checkerboard layout… See the full description on the dataset page: https://huggingface.co/datasets/VatsaDev/cnn_muffins.imageimage-classificationn<1K0 likes106 downloads21d agoHugging Face20Dangindev /VinDR-CXR-VQA VinDr-CXR-VQA Dataset Dataset Description VinDr-CXR-VQA is a large-scale chest X-ray Visual Question Answering (VQA) dataset designed for explainable medical AI with spatial grounding capabilities. The dataset combines natural language question-answer pairs with bounding box annotations and clinical reasoning explanations. Key Features 🏥 4,394 chest X-ray images from VinDr-CXR 💬 17,597 question-answer pairs across 6 question types 📍 Spatial… See the full description on the dataset page: https://huggingface.co/datasets/Dangindev/VinDR-CXR-VQA.textvisual-question-answering1K<n<10K3 likes94 downloads8mo agoHugging Face21shabirahmad5262 /CrisisMMD CrisisMMD: Multimodal Twitter Datasets from Natural Disasters The CrisisMMD multimodal Twitter dataset consists of several thousand manually annotated tweets and images collected during seven major natural disasters, including earthquakes, hurricanes, wildfires, and floods from 2017. The dataset includes three types of annotations: On HuggingFace, we hosted version 2.0 of the CrisisMMD dataset. Please see further information below. Disaster Response Tasks Task 1:… See the full description on the dataset page: https://huggingface.co/datasets/shabirahmad5262/CrisisMMD.imageimage-classification10K<n<100K0 likes86 downloads9mo agoHugging Face22nyuuzyou /OpenGameArt-CC-BY-SA-4.0 Dataset Card for OpenGameArt-CC-BY-SA-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution-ShareAlike 4.0 International (CC-BY-SA-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-SA-4.0.audioimage-classificationn<1K0 likes84 downloads1y agoHugging Face23faizan711 /VinDR-CXR-VQA VinDr-CXR-VQA Dataset Dataset Description VinDr-CXR-VQA is a large-scale chest X-ray Visual Question Answering (VQA) dataset designed for explainable medical AI with spatial grounding capabilities. The dataset combines natural language question-answer pairs with bounding box annotations and clinical reasoning explanations. Key Features 🏥 4,394 chest X-ray images from VinDr-CXR 💬 17,597 question-answer pairs across 6 question types 📍 Spatial grounding with… See the full description on the dataset page: https://huggingface.co/datasets/faizan711/VinDR-CXR-VQA.textvisual-question-answering1K<n<10K3 likes79 downloads8mo agoHugging Face24Daoze /MM-Bench-E-CommerceThis is the HuggingFace repository of the paper named MOON: Generative MLLM-based Multimodal Representation Learning for E-commerce Product Understanding in WSDM 2026 (oral). In this paper, we argue that generative Multimodal Large Language Models (MLLMs) hold significant potential for improving product representation learning. We propose the first generative MLLM-based model named MOON for product representation learning. Furthermore, we contruct and publish a large-scale real-world… See the full description on the dataset page: https://huggingface.co/datasets/Daoze/MM-Bench-E-Commerce.imagetext-classification100K<n<1M2 likes79 downloads7mo agoHugging Face25openlyne /OpenGameArt-CC-BY-4.0 Dataset Card for OpenGameArt-CC-BY-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 4.0 International (CC-BY-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/openlyne/OpenGameArt-CC-BY-4.0.textimage-classification1K<n<10K0 likes71 downloads3mo agoHugging Face26y1665065879 /MOUNT-Cattle Updates/News 📣 🎉 News (Feb. 2026): The dataset paper FSMC-Pose has been accepted for CVPR 2026 Findings! 🔗 News: Please find the open-source dataset on Hugging Face: MOUNT-Cattle. 🔥 Downloads reached 2.4k within 7 days of release. &nbsp;&nbsp; &nbsp;&nbsp; 📌 Overview Mounting posture is an important visual indicator of estrus in dairy cattle. MOUNT-Cattle is a mounting dataset, covering 1,176 mounting… See the full description on the dataset page: https://huggingface.co/datasets/y1665065879/MOUNT-Cattle.imageobject-detection1K<n<10K0 likes63 downloads17d agoHugging Face27nyuuzyou /clker-images Dataset Card for Clker.com Images Dataset Summary This dataset contains 140,313 public domain clipart images collected from Clker.com. Clker.com hosts user-shared vector clip art that is explicitly released into the public domain (CC0). The dataset includes the images themselves along with metadata such as titles and tags associated with each image. Languages The dataset is primarily monolingual: English (en): All image titles and tags are in… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/clker-images.imageimage-classification100K<n<1M2 likes57 downloads1y agoHugging Face28thaotien /movies_CLIP_ViT-L14 🎬 Movie Frame & Caption Dataset 📖 Introduction This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted. This dataset can be used for tasks such as: Video understanding Multimodal learning (image + text) Image captioning Vision-language retrieval 📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.imageimage-classification10K<n<100K0 likes51 downloads1y agoHugging Face29botintel-community /AVAINT-IMGimageimage-to-text100K<n<1M2 likes44 downloads2y agoHugging Face30chenziyue-cattle /MOUNT-Cattle Updates/News 📣 🎉 News (Feb. 2026): The dataset paper FSMC-Pose has been accepted for CVPR 2026 Findings! 🔗 News: Please find the open-source dataset on Hugging Face: MOUNT-Cattle. 🔥 Downloads reached 2.4k within 7 days of release. &nbsp;&nbsp; &nbsp;&nbsp; 📌 Overview Mounting posture is an important visual indicator of estrus in dairy cattle. MOUNT-Cattle is a mounting dataset, covering 1,176 mounting… See the full description on the dataset page: https://huggingface.co/datasets/chenziyue-cattle/MOUNT-Cattle.imageobject-detection1K<n<10K0 likes44 downloads4d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.