datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Anime-Background-Finetuning-V1.1
Anime-Background-Finetuning (10143 manually curated by hand images from danbooru and reddit collections)
The dataset contain roughly 2k of anime Screencap data and 8k of scrapped danbooru illustration data.
This is the proccessed version of the dataset meant to be used for my personal finetuning practice project, please visit my RicemanT/Background-Finetuning repo for the raw unprocessed data that you can process yourself.
The dataset have two minor type of processing being done… See the full description on the dataset page: https://huggingface.co/datasets/RicemanT/Anime-Background-Finetuning-V1.1.image-as-an-imu-finetuning
Image as an IMU: Real-world Finetuning Dataset
Official real-world finetuning dataset from Image as an IMU: Estimating Camera Motion from a Single Motion-Blurred Image (ICCV 2025 Oral).
[arXiv] [Webpage] [GitHub]
PIXL, University of Oxford
Jerred Chen, Ronald Clark
Dataset Details
This dataset consists of 32 sequences of real-world motion-blurred videos in various indoor scenes, captured using the iPhone 13 camera.
dataset_train_real-world.csv and… See the full description on the dataset page: https://huggingface.co/datasets/jerredchen00/image-as-an-imu-finetuning.Anime-Background-Finetuning-V1.1
Anime-Background-Finetuning (10143 manually curated by hand images from danbooru and reddit collections)
The dataset contain roughly 2k of anime Screencap data and 8k of scrapped danbooru illustration data.
This is the proccessed version of the dataset meant to be used for my personal finetuning practice project, please visit my RicemanT/Background-Finetuning repo for the raw unprocessed data that you can process yourself.
The dataset have two minor type of processing being done… See the full description on the dataset page: https://huggingface.co/datasets/HappyHenAi/Anime-Background-Finetuning-V1.1.ddpm-rl-finetuning-evals
Dataset Card for Eval Finetuning Diffusion Models with Reinforcement Learning
XYZ
scanned-images-dataset-for-ocr-and-vlm-finetuning
Dataset Card for scanned_images_dataset
This is a FiftyOne dataset containing 3,482 scanned document images across 10 diverse document categories. Designed for OCR training and Vision-Language Model (VLM) fine-tuning, this dataset features real-world scanned documents with varied layouts, scanning quality, and document types.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/scanned-images-dataset-for-ocr-and-vlm-finetuning.RSVQA-HR_qwen_finetuningFinetuning_Dataset
About:
This dataset is created by Caimera to finetune Diffusion base models to create a finetuned Fashion Diffusion model
brittleness-results
Adapters copied (2026-09-08). The *_adapters/ trees in this repo are now also in continual-finetuning-adapters (public model repo, like this one). Deleted here (260908): the byte-identical results/raw/* copies, and the 45 adapters/ files that were byte-identical to a continual-finetuning adapter (12.3 GB); both lists are in MIGRATION_260908.md of any new repo. Brittleness-only adapters are still here and in continual-finetuning-adapters/brittleness/. Please prefer the new repo for loading.… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/brittleness-results.scanned-images-dataset-for-ocr-and-vlm-finetuning
Dataset Card for scanned_images_dataset
This is a FiftyOne dataset containing 3,482 scanned document images across 10 diverse document categories. Designed for OCR training and Vision-Language Model (VLM) fine-tuning, this dataset features real-world scanned documents with varied layouts, scanning quality, and document types.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from… See the full description on the dataset page: https://huggingface.co/datasets/prabhats0605/scanned-images-dataset-for-ocr-and-vlm-finetuning.hayai-finetuning-dataset-with-koreanAnime_finetuningBD_FinetuningUCMcaptions_finetuningOCR-Finetuning-EN-Dataset
OCR-Finetuning-EN-Dataset
A large-scale English OCR fine-tuning dataset containing synthetic and real-world text images for training modern OCR recognition models.
The dataset is distributed in Apache Parquet format with embedded image data, making it fully compatible with the Hugging Face datasets library and the Hugging Face Dataset Viewer.
Features
✅ 167,330 OCR image-text pairs
✅ Images embedded directly inside Parquet files
✅ Compatible with Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/Srijan-Chakraborty/OCR-Finetuning-EN-Dataset.license-plate-finetuningA formatted, augmented copy of license_plate_object_detection for use with grounding dino training experiments.
Original license is CC - please attribute author at that dataset address.
hayai-finetuning-dataset-final-finalOpenWhistle-Classification-Finetuning
OpenWhistle Classification Finetuning Dataset
OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning is the public
classification finetuning dataset used for dolphin whistle identity
classification. It contains short whistle clips, whistle-level metadata,
fundamental-frequency tracks, rendered F0 spectrograms, and integer class
labels.
The main reviewer-facing subset is the balanced balanced config. It contains
six classes:
NSW_1 (label=0)
SW_Luna (label=1)
SW_Nana (label=2)… See the full description on the dataset page: https://huggingface.co/datasets/OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning.analog_clocks_combinations_for_finetuning
Analog Clocks Combinations Dataset for Finetuning
This repository hosts a collection of 43,200 high-quality, synthetic images of analog clocks, generated for every possible hour, minute, and second in a 12-hour cycle, and for each of three clock types:
Base: normal clocks.
Distorted: dial with distorted shape.
Modified hands: hands with the same thickness and with an arrow.
The data is useful for training and benchmarking computer vision models on tasks like time recognition… See the full description on the dataset page: https://huggingface.co/datasets/migonsa/analog_clocks_combinations_for_finetuning.colpali-finetuning-dataset-gep3
Dataset Card for "colpali-finetuning-dataset-gep3"
More Information needed
colpali-finetuning-dataset-gep2
Dataset Card for "colpali-finetuning-dataset-gep2"
More Information needed
hayai-finetuning-datasetfinetuning_size256_fullfaces100_2split_hflayoutlmv3-finetuning-datafine_tuning_diffusionfinetuning_size256_fullfaces100dreambooth-finetuningfine_tuningdonut_finetuningVLM_FineTuningData_RiskManagementfinetuning_size256_fullfaces100_2split
