datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
countryCounterStrike2Skins
Dataset Card for Counter-Strike 2 Skins Database
Dataset Summary
This dataset contains a comprehensive collection of all skins from Counter-Strike 2. It includes metadata and 1534 high-quality PNG images for each skin. The dataset is useful for researchers, developers, building applications related to CS2 skins.
Dataset Structure
Data Format
The dataset is provided in JSON format, where each entry represents a skin with associated metadata:
{… See the full description on the dataset page: https://huggingface.co/datasets/While402/CounterStrike2Skins.CountBenchQAThis dataset was introduced in PaliGemma for evaluating counting in vision language models. This version only includes 491 images from the original CountBench dataset, since some of the original URLs can no longer be accessed.
Original Description
CountBench: We introduce a new object counting benchmark called CountBench,
automatically curated (and manually verified) from the publicly available
LAION-400M image-text dataset. CountBench contains a total of 540 images
containing… See the full description on the dataset page: https://huggingface.co/datasets/vikhyatk/CountBenchQA.CountQA
Dataset Summary
CountQA is the new benchmark designed to stress-test the Achilles' heel of even the most advanced Multimodal Large Language Models (MLLMs): object counting. While modern AI demonstrates stunning visual fluency, it often fails at this fundamental cognitive skill, a critical blind spot limiting its real-world reliability.
This dataset directly confronts that weakness with over 1,500 challenging question-answer pairs built on real-world images, hand-captured to feature… See the full description on the dataset page: https://huggingface.co/datasets/Jayant-Sravan/CountQA.google-streetview-images-by-country
Dataset Card for google streetview images by country
⚠️ There are still images that should be deleted, such as those with tags or those that didn't load correctly.
Dataset Structure
folder with the individual countries
images have the creation date and the map name in the file name.
Dataset Card Contact
use the community section
images per country
GeoGuessr-countries-largepixmo-count
PixMo-Count
PixMo-Count is a dataset of images paired with objects and their point locations in the image.
It was built by running the Detic object detector on web images, and then filtering the data
to improve accuracy and diversity. The val and test sets are human-verified and only contain counts from 2 to 10.
PixMo-Count is a part of the PixMo dataset collection and was used to
augment the pointing capabilities of the Molmo family of models
Quick links:
📃 Paper
🎥 Blog with… See the full description on the dataset page: https://huggingface.co/datasets/allenai/pixmo-count.ai2thor-counting-largepixmo-point-count-concat_0-20Anon-CounterFactual-Dataset
Anon Counterfactual Dataset
Dataset repo: https://huggingface.co/datasets/dataset-author-404/Anon-CounterFactual-Dataset
Synthetic CLEVR-style 3D scenes with original, semantic counterfactual, and negative (artifact) PNG renders, plus VQA-style questions, difficulties, and a 3×3 answer matrix. Built from the MMB counterfactual pipeline run folder dataset_720p_v2 (see build_hub_dataset.py in this repo snapshot).
This revision replaces the previous Hub layout (legacy imagefolder /… See the full description on the dataset page: https://huggingface.co/datasets/dataset-author-404/Anon-CounterFactual-Dataset.procthor-100-counting-balancedMulti-Hop-Objects-Countingpixmo-point-count-gen-undwds_country211country211
Dataset Card for Country211
The Country 211 Dataset from OpenAI.
This dataset was built by filtering the images from the YFCC100m dataset that have GPS coordinate corresponding to a ISO-3166 country code. The dataset is balanced by sampling 150 train images, 50 validation images, and 100 test images images for each country.
crowd-counting
Crowd Density Dataset - Different Crowd Sizes
The dataset consists of 647 images of crowds, containing up to 11,000 individuals, annotated with keypoints for precise crowd counting and density estimation. It is designed for crowd counting tasks, particularly in crowded scenes settings, accommodating various sizes and challenges in estimating density. The dataset includes examples of both denser crowds and sparser crowds, enhancing counting accuracy for real-world applications in… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/crowd-counting.worldcuisines_format_sea_country_only_with_metadatacountbench
Dataset Card for "countbench"
This dataset was introduced in the paper Teaching CLIP to Count to Ten.
counterfactual-physicssd-prompt-image-in-the-wild-counterfeitwds_vtab-clevr_count_allVisual-Counterfact
Visual CounterFact: Controlling Knowledge Priors in Vision-Language Models through Visual Counterfactuals
This dataset is part of the work "Pixels Versus Priors: Controlling Knowledge Priors in Vision-Language Models through Visual Counterfacts".📖 Read the Paper💾 GitHub Repository
Overview
Visual CounterFact is a novel dataset designed to investigate how Multimodal Large Language Models (MLLMs) balance memorized world knowledge priors (e.g., "strawberries are red")… See the full description on the dataset page: https://huggingface.co/datasets/mgolov/Visual-Counterfact.clevr_count_70kThis dataset is borrowed from clevr_cogen_a_train
ai2thor-multiview-counting-val-800-v2shanghaitech-crowd-countingcounterfactual-pendulum-multilingual
📌 Dataset Summary
When a Vision-Language Model (VLM) is given an image along with a text prompt containing contradictory or misleading information, how does it react? Does it rely on the visual evidence, succumb to textual bias, or honestly abstain when faced with unresolvable conflict?
This dataset adapts the Counterfactual Pendulum scenario across two visual conflict dimensions:
Angular (Angle): Conflict in the pendulum's angle of inclination.
Light: Conflict in the light… See the full description on the dataset page: https://huggingface.co/datasets/apart-global-south-hack/counterfactual-pendulum-multilingual.COUNTSai2thor-counting-final-400MTDC_Maize_Tassels_Detection_Counting_Dataset
MTDC — Maize Tassels Detection and Counting Dataset
A UAV-imagery dataset for maize tassel detection and counting on maize crops. The dataset contains 361 images with 13,564 annotated maize tassel instances (single class: maize). Bounding boxes are in COCO format [x, y, width, height].
Split
Images
Annotations
train
361
13,564
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.
Loading
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/MTDC_Maize_Tassels_Detection_Counting_Dataset.coco-counterfactual-conflict
COCO-Counterfactual Conflict
Image-text conflict dataset built from
Intel/COCO-Counterfactuals.
Each COCO-Counterfactuals example is a minimal pair of captions differing by a single noun
subject, with a matching image for each. We keep the truthful image (image_0) and its
caption as original_caption, and use the counterfactual caption as conflicting_caption.
The swapped noun is extracted automatically (image_bias = true noun, text_bias = altered
noun); the question and… See the full description on the dataset page: https://huggingface.co/datasets/multilingual-vlm-conflict/coco-counterfactual-conflict.
