CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01vikhyatk /CountBenchQAThis dataset was introduced in PaliGemma for evaluating counting in vision language models. This version only includes 491 images from the original CountBench dataset, since some of the original URLs can no longer be accessed. Original Description CountBench: We introduce a new object counting benchmark called CountBench, automatically curated (and manually verified) from the publicly available LAION-400M image-text dataset. CountBench contains a total of 540 images containing… See the full description on the dataset page: https://huggingface.co/datasets/vikhyatk/CountBenchQA.imagen<1K9 likes3.4k downloads2y agoHugging Face02Jayant-Sravan /CountQA Dataset Summary CountQA is the new benchmark designed to stress-test the Achilles' heel of even the most advanced Multimodal Large Language Models (MLLMs): object counting. While modern AI demonstrates stunning visual fluency, it often fails at this fundamental cognitive skill, a critical blind spot limiting its real-world reliability. This dataset directly confronts that weakness with over 1,500 challenging question-answer pairs built on real-world images, hand-captured to feature… See the full description on the dataset page: https://huggingface.co/datasets/Jayant-Sravan/CountQA.imagevisual-question-answering1K<n<10K5 likes2.3k downloads1y agoHugging Face03allenai /pixmo-count PixMo-Count PixMo-Count is a dataset of images paired with objects and their point locations in the image. It was built by running the Detic object detector on web images, and then filtering the data to improve accuracy and diversity. The val and test sets are human-verified and only contain counts from 2 to 10. PixMo-Count is a part of the PixMo dataset collection and was used to augment the pointing capabilities of the Molmo family of models Quick links: 📃 Paper 🎥 Blog with… See the full description on the dataset page: https://huggingface.co/datasets/allenai/pixmo-count.imagevisual-question-answering10K<n<100K12 likes860 downloads2y agoHugging Face04Jiwon-Kang /pixmo-point-count-concat_0-20image100K<n<1M0 likes578 downloads9mo agoHugging Face05multimodal-reasoning-lab /Multi-Hop-Objects-Countingimage10K<n<100K5 likes525 downloads1y agoHugging Face06heez /pixmo-point-count-gen-undimage100K<n<1M0 likes483 downloads8mo agoHugging Face07nateraw /country211 Dataset Card for Country211 The Country 211 Dataset from OpenAI. This dataset was built by filtering the images from the YFCC100m dataset that have GPS coordinate corresponding to a ISO-3166 country code. The dataset is balanced by sampling 150 train images, 50 validation images, and 100 test images images for each country. imageimage-classification10K<n<100K6 likes440 downloads4y agoHugging Face08nielsr /countbench Dataset Card for "countbench" This dataset was introduced in the paper Teaching CLIP to Count to Ten. imagen<1K10 likes384 downloads6mo agoHugging Face09la-ji /sd-prompt-image-in-the-wild-counterfeitimage1M<n<10M3 likes335 downloads2y agoHugging Face10SEACrowd /worldcuisines_format_sea_country_only_with_metadataimage100K<n<1M0 likes296 downloads10mo agoHugging Face11phy-gen /counterfactual-physicsimage10K<n<100K0 likes296 downloads3mo agoHugging Face12Jiwon-Kang /pixmo-count-filtered-imgContainedimage10K<n<100K0 likes270 downloads9mo agoHugging Face13BUAADreamer /clevr_count_70kThis dataset is borrowed from clevr_cogen_a_train image10K<n<100K3 likes233 downloads2y agoHugging Face14weikaih /ai2thor-counting-largeimage1K<n<10K1 likes216 downloads1y agoHugging Face15mgolov /Visual-Counterfact Visual CounterFact: Controlling Knowledge Priors in Vision-Language Models through Visual Counterfactuals This dataset is part of the work "Pixels Versus Priors: Controlling Knowledge Priors in Vision-Language Models through Visual Counterfacts".📖 Read the Paper💾 GitHub Repository Overview Visual CounterFact is a novel dataset designed to investigate how Multimodal Large Language Models (MLLMs) balance memorized world knowledge priors (e.g., "strawberries are red")… See the full description on the dataset page: https://huggingface.co/datasets/mgolov/Visual-Counterfact.imageimage-text-to-text1K<n<10K4 likes209 downloads1y agoHugging Face16jxie /country211 Dataset Card for "country211" More Information needed image10K<n<100K0 likes206 downloads3y agoHugging Face17KTAEHWA /shanghaitech-crowd-countingimage1K<n<10K1 likes189 downloads10mo agoHugging Face18Gigagiggles /european-countries-classifierimage1K<n<10K0 likes168 downloads9mo agoHugging Face19leo66666 /scannet_countingimagen<1K0 likes143 downloads7mo agoHugging Face20multilingual-vlm-conflict /coco-counterfactual-conflict COCO-Counterfactual Conflict Image-text conflict dataset built from Intel/COCO-Counterfactuals. Each COCO-Counterfactuals example is a minimal pair of captions differing by a single noun subject, with a matching image for each. We keep the truthful image (image_0) and its caption as original_caption, and use the counterfactual caption as conflicting_caption. The swapped noun is extracted automatically (image_bias = true noun, text_bias = altered noun); the question and… See the full description on the dataset page: https://huggingface.co/datasets/multilingual-vlm-conflict/coco-counterfactual-conflict.imagevisual-question-answering1K<n<10K0 likes137 downloads3mo agoHugging Face21dnth /pixmo-count-imagesimage1K<n<10K0 likes125 downloads2y agoHugging Face22ShyFoo /CountHallu-Dataset-SimObject CountHalluSet — SimObject Rendered dataset from Counting Hallucinations in Diffusion Models (arXiv:2510.13080). Part of CountHalluSet, a suite with well-defined counting criteria used to measure counting hallucination — a diffusion model generating the wrong number of instances, even for patterns absent from its training data. What's inside 256×256 RGB rendered images of everyday objects, each labelled with the per-class instance count over three object classes.… See the full description on the dataset page: https://huggingface.co/datasets/ShyFoo/CountHallu-Dataset-SimObject.imageunconditional-image-generation10K<n<100K1 likes114 downloads2mo agoHugging Face23gsarch /countqa_lite gsarch/countqa_lite A deterministic lite evaluation subset of Jayant-Sravan/CountQA. Source revision: f92cc6fe46542c61e2916e3d2ae9a911e2216b1a Source split: test Sampling seed: 43 Output rows: 500 Schema: unchanged from the upstream dataset CountQA is sampled at the QA-pair level. Each output row retains the original schema and contains one-element questions and answers lists, so lmms-eval's existing countqa_process_docs produces exactly 500 prompts. Generated by… See the full description on the dataset page: https://huggingface.co/datasets/gsarch/countqa_lite.imagevisual-question-answeringn<1K0 likes113 downloads2mo agoHugging Face24dddraxxx /single-plot-count-3kimage1K<n<10K0 likes105 downloads1y agoHugging Face25TMLR-Group-HF /counteranimalimage10K<n<100K1 likes98 downloads1y agoHugging Face26tackhwa /worldcuisines_format_sea_country_only_3image100K<n<1M0 likes96 downloads10mo agoHugging Face27surogate /ro_sft_pixmo_count Dataset Description PixmoCount is a dataset of images paired with number of objects in the image. Here we provide the Romanian translation of the PixmoCount dataset, translated with Seed-X-PPO. This dataset is part of the instruction finetune protocol for Romanian VLMs proposed in "Înțelegi românește?" A Recipe for Romanian Vision-Language Models (Masala et al., 2026). Citation @inproceedings{deitke2025molmo, title={Molmo and pixmo: Open weights and open data… See the full description on the dataset page: https://huggingface.co/datasets/surogate/ro_sft_pixmo_count.image10K<n<100K0 likes95 downloads1mo agoHugging Face28ioaihsc /Task2_Chicken_Counting_Train2imagen<1K0 likes78 downloads1y agoHugging Face29dll-streetview /sem-seg-country-safety-binsimage10K<n<100K0 likes71 downloads2y agoHugging Face30Aniket96 /country-flags-dataset World Country Flags Dataset This dataset contains flag images from sovereign states with their country names as labels. Dataset Description A comprehensive collection of flag images for all sovereign nations, organized for machine learning tasks. Dataset Summary Total Images: 195 country flags Format: PNG (640x427 pixels) Task: Image classification Language: English country names Dataset Structure Each example contains: image: The flag image in… See the full description on the dataset page: https://huggingface.co/datasets/Aniket96/country-flags-dataset.imageimage-classificationn<1K2 likes70 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.