datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ViT-FineTunemovies_CLIP_ViT-L14
🎬 Movie Frame & Caption Dataset
📖 Introduction
This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted.
This dataset can be used for tasks such as:
Video understanding
Multimodal learning (image + text)
Image captioning
Vision-language retrieval
📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.index-cards-southborough-vital-records
Southborough Town Clerk — Vital Records & Veteran Index Cards (MA)
8,661 cards from the Southborough (Massachusetts) Town Clerk office, covering:
Death Index Cards, 1850–2015 — the town clerk's running death index covering ~165 years of Southborough deaths, one card per decedent with surname/given-name/date.
Veteran Card Index + Veteran Grave Registration Card Index — companion indices to Southborough's veteran-affairs records, indexing veterans buried in town cemeteries.… See the full description on the dataset page: https://huggingface.co/datasets/biglam/index-cards-southborough-vital-records.waste-segregation-vit
Waste Segregation Dataset (ViT Fine-tuning)
Descripción
Dataset de clasificación de residuos adaptado para el afinamiento de modelos Vision Transformer (ViT).
Contiene imágenes de diferentes categorías de desechos, divididas en conjuntos de entrenamiento,
validación y prueba siguiendo las convenciones de Hugging Face Datasets.
Fuente Original
Kaggle: Waste Segregation – smarthkaushal
Estructura
Split
Ejemplos
train
~70%
validation
~15%… See the full description on the dataset page: https://huggingface.co/datasets/Dagachaco/waste-segregation-vit.
