datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
japanese-aerial-fireworks-v2
🎆 NEW: Curated 1,000 Wide Pack (Commercial License)
For commercial AI/ML training, check out the Hanabi AI Dataset v1: Wide Pack — Curated 1,000 — a carefully selected subset with detailed structured annotations:
✅ 1,000 hand-curated 4K images (vs 2,557 raw images here)
✅ Structured AI annotations (composition, mood, color, EXIF, English notes)
✅ Sample PyTorch loader, attribute filter, caption generator
✅ Perpetual Commercial License (Japanese law)
✅ Optimized for Stable… See the full description on the dataset page: https://huggingface.co/datasets/dfhjs2577/japanese-aerial-fireworks-v2.FireBench
FireBench: A Benchmark Dataset for Fire Science Image Retrieval
Dataset Description
FireBench is a benchmark dataset for evaluating image retrieval systems in the domain of wildfire and fire science. The dataset consists of natural language queries paired with images, along with binary relevance labels indicating whether each image is relevant to the query. The dataset is designed to test retrieval systems' ability to find relevant wildfire-related images based on a… See the full description on the dataset page: https://huggingface.co/datasets/sagecontinuum/FireBench.
