datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CommonSketch
CommonSketch
Dataset Summary
CommonSketch is a semantically annotated sketch dataset introduced in the paper SEA: Evaluating Sketch Abstraction Efficiency via Element-level Commonsense Visual Question Answering. The dataset contains 23,100 human-drawn sketches across 300 object classes. Each sketch is paired with a fine-grained caption and element-level commonsense annotations for evaluating sketch abstraction and semantic recognizability.
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/ziiio/CommonSketch.CommonObjectsBench
CommonObjectsBench: A Benchmark Dataset for General Object Image Retrieval
Dataset Description
CommonObjectsBench is a benchmark dataset for evaluating image retrieval systems on general objects and common scenes. The dataset consists of natural language queries paired with images, along with binary relevance labels indicating whether each image is relevant to the query. The dataset is designed to test retrieval systems' ability to find relevant images based on queries… See the full description on the dataset page: https://huggingface.co/datasets/sagecontinuum/CommonObjectsBench.University_CommonLostItemsflickr-lifeboat-commons-1k-2025
Flickr Commons 1K Collection
Dataset Description
This dataset is a Flickr Data Lifeboat converted to a machine learning ready/ Hugging Face datasets compatible format. Data Lifeboats are digital preservation archives created by the Flickr Foundation to ensure long-term access to meaningful collections of Flickr photos and their rich community metadata.
What is a Data Lifeboat?
Data Lifeboats are self-contained archives designed to preserve not just images… See the full description on the dataset page: https://huggingface.co/datasets/flickr-foundation/flickr-lifeboat-commons-1k-2025.
