datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
google-streetview-images-by-country
Dataset Card for google streetview images by country
⚠️ There are still images that should be deleted, such as those with tags or those that didn't load correctly.
Dataset Structure
folder with the individual countries
images have the creation date and the map name in the file name.
Dataset Card Contact
use the community section
images per country
country211
Dataset Card for Country211
The Country 211 Dataset from OpenAI.
This dataset was built by filtering the images from the YFCC100m dataset that have GPS coordinate corresponding to a ISO-3166 country code. The dataset is balanced by sampling 150 train images, 50 validation images, and 100 test images images for each country.
crowd-counting
Crowd Density Dataset - Different Crowd Sizes
The dataset consists of 647 images of crowds, containing up to 11,000 individuals, annotated with keypoints for precise crowd counting and density estimation. It is designed for crowd counting tasks, particularly in crowded scenes settings, accommodating various sizes and challenges in estimating density. The dataset includes examples of both denser crowds and sparser crowds, enhancing counting accuracy for real-world applications in… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/crowd-counting.CountHallu-Dataset-SimObject
CountHalluSet — SimObject
Rendered dataset from Counting Hallucinations in Diffusion Models
(arXiv:2510.13080). Part of CountHalluSet, a suite with well-defined counting
criteria used to measure counting hallucination — a diffusion model generating
the wrong number of instances, even for patterns absent from its training data.
What's inside
256×256 RGB rendered images of everyday objects, each labelled with the
per-class instance count over three object classes.… See the full description on the dataset page: https://huggingface.co/datasets/ShyFoo/CountHallu-Dataset-SimObject.geoguessr-countries-finetune
GeoGuessr Countries Finetune
Google Earth images from around the world with the country as the target label.
Splits
Split
Samples
train
25000
test
400
Columns
image: Google Earth image
country: country label
Countries In This Release
Argentina, Australia, Austria, Bangladesh, Belgium, Bolivia, Botswana, Brazil, Bulgaria,
Cambodia, Canada, Chile, Colombia, Croatia, Czechia, Denmark, Finland, France, Germany,
Ghana, Greece, Hungary… See the full description on the dataset page: https://huggingface.co/datasets/moondream/geoguessr-countries-finetune.country-flags-dataset
World Country Flags Dataset
This dataset contains flag images from sovereign states with their country names as labels.
Dataset Description
A comprehensive collection of flag images for all sovereign nations, organized for machine learning tasks.
Dataset Summary
Total Images: 195 country flags
Format: PNG (640x427 pixels)
Task: Image classification
Language: English country names
Dataset Structure
Each example contains:
image: The flag image in… See the full description on the dataset page: https://huggingface.co/datasets/Aniket96/country-flags-dataset.CountHallu-dataset-ToyShape
CountHalluSet — ToyShape
Synthetic dataset from Counting Hallucinations in Diffusion Models
(arXiv:2510.13080). Part of CountHalluSet, a suite with well-defined counting
criteria used to measure counting hallucination — a diffusion model generating
the wrong number of instances, even for patterns absent from its training data.
What's inside
128×128 RGB images of non-overlapping white shapes on a black background. Each
image holds 1–3 shapes drawn from {triangle… See the full description on the dataset page: https://huggingface.co/datasets/ShyFoo/CountHallu-dataset-ToyShape.shape-counting-dataset
Shape Counting Dataset
A dataset for evaluating shape counting abilities in vision models and humans.
Dataset Description
This dataset contains images with varying numbers of squares, triangles, and stars on a white background. Each image is provided in multiple versions: the original clean image plus several noisy variants.
Image Specifications
Size: 256×256 pixels
Format: Grayscale PNG
Shape size: 18 pixels
Background: White (255)
Shapes: Black (0)… See the full description on the dataset page: https://huggingface.co/datasets/nooranis/shape-counting-dataset.Crowd-Countin-Dataset
Crowd Dataset - 647 Photos
Dataset comprises 647 photos of dense crowds, containing between 1,000 to 13,000 people per image. Each image includes detailed keypoint annotations for every individual, enabling advanced data analysis and deep learning applications in crowd density estimation, object detection, and counting algorithms. - Get the data
Dataset characteristics:
Characteristic
Data
Description
Crowd photos with labeling for determining crowd density… See the full description on the dataset page: https://huggingface.co/datasets/ud-smart-city/Crowd-Countin-Dataset.country-flags-variations
Country Flag Variations (Flux 2-dev, image-verified)
Synthetic flag images generated with Flux 2-dev from the prompt
"The national flag of {country}", then filtered by a strict
image-to-image verification against official reference flags.
Dataset summary
71 countries, 1395 images
Each country has 15-20 verified images (originally 20 per country;
images that disagreed with the reference flag were removed)
Split: per-country 80 / 20 train/val (deterministic by sorted seed)… See the full description on the dataset page: https://huggingface.co/datasets/tokeron/country-flags-variations.autotrain-data-square-count-classifier
AutoTrain Dataset for project: square-count-classifier
Dataset Description
This dataset has been automatically processed by AutoTrain for project square-count-classifier.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<28x28 L PIL image>",
"target": 0
},
{
"image": "<28x28 L PIL image>",
"target": 0
}
]… See the full description on the dataset page: https://huggingface.co/datasets/dhavala/autotrain-data-square-count-classifier.country-flags-dataset
World Country Flags Dataset
This dataset contains flag images from sovereign states with their country names as labels.
Dataset Description
A comprehensive collection of flag images for all sovereign nations, organized for machine learning tasks.
Dataset Summary
Total Images: 195 country flags
Format: PNG (640x427 pixels)
Task: Image classification
Language: English country names
Dataset Structure
Each example contains:
image: The flag image in… See the full description on the dataset page: https://huggingface.co/datasets/bigggggerds/country-flags-dataset.country-flags-dataset
World Country Flags Dataset
This dataset contains flag images from sovereign states with their country names as labels.
Dataset Description
A comprehensive collection of flag images for all sovereign nations, organized for machine learning tasks.
Dataset Summary
Total Images: 195 country flags
Format: PNG (640x427 pixels)
Task: Image classification
Language: English country names
Dataset Structure
Each example contains:
image: The… See the full description on the dataset page: https://huggingface.co/datasets/logan3027/country-flags-dataset.
