CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Shiym /ViT-FineTuneimageimage-classification10K<n<100K0 likes5.2k downloads2y agoHugging Face02thaotien /movies_CLIP_ViT-L14 🎬 Movie Frame & Caption Dataset 📖 Introduction This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted. This dataset can be used for tasks such as: Video understanding Multimodal learning (image + text) Image captioning Vision-language retrieval 📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.imageimage-classification10K<n<100K0 likes38 downloads1y agoHugging Face03biglam /index-cards-southborough-vital-records Southborough Town Clerk — Vital Records & Veteran Index Cards (MA) 8,661 cards from the Southborough (Massachusetts) Town Clerk office, covering: Death Index Cards, 1850–2015 — the town clerk's running death index covering ~165 years of Southborough deaths, one card per decedent with surname/given-name/date. Veteran Card Index + Veteran Grave Registration Card Index — companion indices to Southborough's veteran-affairs records, indexing veterans buried in town cemeteries.… See the full description on the dataset page: https://huggingface.co/datasets/biglam/index-cards-southborough-vital-records.imageimage-to-text1K<n<10K0 likes24 downloads4mo agoHugging Face04s-emanuilov /coco-clip-vit-l-14 COCO Dataset Processed with CLIP ViT-L/14 Overview This dataset represents a processed version of the '2017 Unlabeled images' subset of the COCO dataset (COCO Dataset), utilizing the CLIP ViT-L/14 model from OpenAI. The original dataset comprises 123K images, approximately 19GB in size, which have been processed to generate 786-dimensional vectors. These vectors can be utilized for various applications like semantic search systems, image similarity assessments, and more.… See the full description on the dataset page: https://huggingface.co/datasets/s-emanuilov/coco-clip-vit-l-14.textimage-classification100K<n<1M2 likes17 downloads2y agoHugging Face05roth1414 /galaxy-vit-gz-desi-dirichlet-predictions Galaxy-ViT — GZ DESI Dirichlet predictions Per-galaxy Dirichlet-Multinomial concentration parameters (α) for the 10-question / 34-answer Galaxy Zoo DESI decision tree, predicted by a Zoobot ConvNeXt-nano encoder finetuned with a Dirichlet-Multinomial head on the DR8 subset of the mwalmsley/gz_desi_wds labeled split. Dataset summary Rows 61,440 Columns 36 (key, dr8_id, alpha_0 … alpha_33) Format Apache Parquet File size ~16 MB Source images DECaLS DR8… See the full description on the dataset page: https://huggingface.co/datasets/roth1414/galaxy-vit-gz-desi-dirichlet-predictions.tabulartabular-classification10K<n<100K0 likes14 downloads5mo agoHugging Face06Dagachaco /waste-segregation-vit Waste Segregation Dataset (ViT Fine-tuning) Descripción Dataset de clasificación de residuos adaptado para el afinamiento de modelos Vision Transformer (ViT). Contiene imágenes de diferentes categorías de desechos, divididas en conjuntos de entrenamiento, validación y prueba siguiendo las convenciones de Hugging Face Datasets. Fuente Original Kaggle: Waste Segregation – smarthkaushal Estructura Split Ejemplos train ~70% validation ~15%… See the full description on the dataset page: https://huggingface.co/datasets/Dagachaco/waste-segregation-vit.imageimage-classification1K<n<10K0 likes9 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.