datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
saikyounoousamanidomenojinseiwananiwosuru
Bangumi Image Base of Saikyou No Ousama, Nidome No Jinsei Wa Nani Wo Suru?
This is the image base of bangumi Saikyou no Ousama, Nidome no Jinsei wa Nani wo Suru?, we detected 71 characters, 4913 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saikyounoousamanidomenojinseiwananiwosuru.sailormoon1990s
Bangumi Image Base of Sailor Moon (1990s)
This is the image base of bangumi Sailor Moon (1990s), we detected 132 characters, 14684 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/sailormoon1990s.SAINetset_v8.0
SAINetset - Wildfire Smoke Detection Dataset
Dataset of real-world images captured by SAI (Sistema de Alerta de Incendios / Fire Alert System) surveillance nodes for wildfire smoke detection in Cordoba, Argentina.
Current version: v8.0 (January 2026)
About SAI
The SAI (Fire Alert System) is an open-source early wildfire detection platform developed by AlterMundi, a civil association in Argentina. The system uses distributed camera nodes with YOLO-based AI (powered by… See the full description on the dataset page: https://huggingface.co/datasets/SAINetset/SAINetset_v8.0.gs-lrmspinetrackPlease visit the project homepage for more details about the dataset, associated research paper, and citation information.
If you use our dataset in your work, we request proper attribution and a citation to our paper "Towards Unconstrained 2D Pose Estimation of the Human Spine".
saikyounoshienshokuwajutsushidearuorewasekaisaikyouclanwoshitagaeru
Bangumi Image Base of Saikyou No Shienshoku "wajutsushi" De Aru Ore Wa Sekai Saikyou Clan Wo Shitagaeru
This is the image base of bangumi Saikyou no Shienshoku "Wajutsushi" de Aru Ore wa Sekai Saikyou Clan wo Shitagaeru, we detected 70 characters, 4558 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saikyounoshienshokuwajutsushidearuorewasekaisaikyouclanwoshitagaeru.sainfoin-seed-datasetsaikyoutanknomeikyuukouryaku
Bangumi Image Base of Saikyou Tank No Meikyuu Kouryaku
This is the image base of bangumi Saikyou Tank no Meikyuu Kouryaku, we detected 86 characters, 5885 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1%… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saikyoutanknomeikyuukouryaku.saijakutamerwagomihiroinotabiwohajimemashita
Bangumi Image Base of Saijaku Tamer Wa Gomi Hiroi No Tabi Wo Hajimemashita
This is the image base of bangumi Saijaku Tamer wa Gomi Hiroi no Tabi wo Hajimemashita, we detected 81 characters, 6058 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saijakutamerwagomihiroinotabiwohajimemashita.saihatenopaladin
Bangumi Image Base of Saihate No Paladin
This is the image base of bangumi Saihate no Paladin, we detected 73 characters, 7532 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is the… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saihatenopaladin.sailormoon2010s
Bangumi Image Base of Sailor Moon (2010s)
This is the image base of bangumi Sailor Moon (2010s), we detected 46 characters, 3463 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).
Here is… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/sailormoon2010s.unified-vlm-steering-emu35
Emu3.5 (BAAI): activation steering sweeps
Steered text and image generations from Emu3.5 (BAAI), one of the unified vision-language models in the unified-vlm-steering project. Steering adds alpha * v_hat (the per-layer unit difference-of-means vector) to the residual stream at every layer of a layer config.
34B dense autoregressive model, 64 decoder layers (0-indexed). Images are 32x32 IBQ tokens (512 px), generated on BAAI's patched vLLM engine with request-id-keyed CFG. The… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/unified-vlm-steering-emu35.unified-vlm-steering-uniar
UniAR (ShareLab-SII/UniAR-RL): activation steering sweeps
Steered text and image generations from UniAR (ShareLab-SII/UniAR-RL), one of the unified vision-language models in the unified-vlm-steering project. Steering adds alpha * v_hat (the per-layer unit difference-of-means vector) to the residual stream at every layer of a layer config.
Qwen3-VL backbone, 36 decoder layers (0-indexed). Images are BSQ tokens rendered by an SD3 decoder (16 decoding steps, 544 px).… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/unified-vlm-steering-uniar.saikinyatottamaidgaayashii
Bangumi Image Base of Saikin Yatotta Maid Ga Ayashii
This is the image base of bangumi Saikin Yatotta Maid ga Ayashii, we detected 26 characters, 4157 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1%… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saikinyatottamaidgaayashii.unified-vlm-steering-liquid
Liquid (FoundationVision Liquid_V1_7B): activation steering sweeps
Steered text and image generations from Liquid (FoundationVision Liquid_V1_7B), one of the unified vision-language models in the unified-vlm-steering project. Steering adds alpha * v_hat (the per-layer unit difference-of-means vector) to the residual stream at every layer of a layer config.
Gemma-7B backbone, 28 decoder layers (0-indexed). Images are VQGAN codes (512 px), CFG 7.0, top-k 4096, top-p 0.96… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/unified-vlm-steering-liquid.benchmark_omni3DOMNI Benchmark Dataset Description and Metadata
saikyouonmyoujinoisekaitenseiki
Bangumi Image Base of Saikyou Onmyouji No Isekai Tenseiki
This is the image base of bangumi Saikyou Onmyouji no Isekai Tenseiki, we detected 63 characters, 5099 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saikyouonmyoujinoisekaitenseiki.ArtVee_datasetsaihatenopaladintetsusabinoyamanoou
Bangumi Image Base of Saihate No Paladin: Tetsusabi No Yama No Ou
This is the image base of bangumi Saihate no Paladin: Tetsusabi no Yama no Ou, we detected 35 characters, 3323 images in total. The full dataset is here.
Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/saihatenopaladintetsusabinoyamanoou.car-images
Dataset Card for Car Images Dataset
Dataset Description
This dataset contains a collection of car images designed to support tasks such as image classification, car detection, and autonomous driving research. The images feature cars in various settings, including streets and other environments, captured to provide diverse training data for computer vision models.
The dataset aims to enable the development and benchmarking of models that can:
Identify… See the full description on the dataset page: https://huggingface.co/datasets/saifurrehman1234567/car-images.uniar-steering-eval
UniAR Steering Eval — text quadrants (samples only)
Activation-steering text generations from UniAR (Qwen3-VL backbone, BSQ visual tokens, SD3 decoder).
Raw generated samples only — the LLM-judge scores have been removed.
Quadrants (2, text output)
A steering direction is extracted from minimal pairs in one modality, then injected during pure text generation.
sub
vector from
steers
output
txt2txt
text pairs
whole user message + generation
text
img2txt… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/uniar-steering-eval.benchmark_coco_filteredCOCO Benchmark Dataset Description and Metadata
unified-vlm-steering-minimal-pairs
Unified VLM Steering: Minimal Pairs
Minimal pairs used to extract steering vectors (difference of means between the two poles) for 7 concepts in the
unified-vlm-steering project: 100 text pairs and 100 image pairs per concept.
Layout
txt/<concept>/pairs.json 100 text pairs: concept, pos_label, neg_label, template, n_pairs,
pairs [{subject, pos, neg}], and the pos / neg sentence lists
img/<concept>/<000-099>/
baseline.png… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/unified-vlm-steering-minimal-pairs.safemaize-screening3-v3
SafeMaize screening3_v3
Public maize images for screening experiments on northern/turcicum leaf blight-like
symptoms and fall armyworm, with healthy maize as the comparison. Images keep their
original bytes and are stored as Parquet shards, so no loose image files are published.
Class
Core images
healthy
29,448
nlb_tlb_like
23,889
faw_feeding_injury
10,459
Partition
Images
Shards
train
43,945
74
val
5,669
13
calibration
5,669
13
test
8,513
11… See the full description on the dataset page: https://huggingface.co/datasets/saiteja33/safemaize-screening3-v3.unified-vlm-steering-eval
Unified VLM steering eval
Activation steering of unified vision-language models: steered text and image generations, steering vectors and
judge scores. One folder per model; prompts/ is shared. Steering is h <- h + alpha * v_hat everywhere.
model
run
text rows
image rows
concepts with images
published
uniar/
av_v2
21,560
21,560
age, chaos, cleanness, emotion, near_far, size, spatial_lr
2026-09-11
Layout per model: manifest.json (exact config)… See the full description on the dataset page: https://huggingface.co/datasets/saintsauce/unified-vlm-steering-eval.StomataBenchtea-leaf-disease-dataset
Data
Image data is not versioned in this repository. Only manifests, audit outputs,
and fingerprints are tracked, which is enough to verify and rebuild the exact
partition used.
Layout
data/
raw/ TeaLeafBD release as downloaded (gitignored)
tea_leaf_clean/ intermediate build (gitignored)
Tea_leaf_dataset/ FINAL audited dataset (gitignored)
manifests/ tracked audit artifacts
ANTroof-top-dataset
🏠 Roof Top Segmentation Dataset
A dataset for roof segmentation and edge detection from satellite imagery,
with annotations derived from AutoCAD DXF drawings.
Dataset Columns
Each row is one roof sample:
Column
Type
Description
sample_id
string
Unique ID (e.g. roof_001)
image
Image
Satellite roof photograph
mask
Image
Binary segmentation mask (white=roof)
overlay
Image
Verification overlay with edge lines
outer_polygon
string (JSON)
Roof boundary… See the full description on the dataset page: https://huggingface.co/datasets/Sai150/roof-top-dataset.American-Sign-Language-MNIST
Dataset Card for ASL-MNIST
This is a FiftyOne dataset with 34,627 samples of American Sign Language (ASL) alphabet images, converted from the original Kaggle Sign Language MNIST dataset into a format optimized for computer vision workflows.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available… See the full description on the dataset page: https://huggingface.co/datasets/saifmughal23/American-Sign-Language-MNIST.
