CoolFace
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ego-thales /cifar10 Dataset Specifications Contains the entire CIFAR10 dataset, downloaded via PyTorch, then split and saved as .png files representing 32x32 images. There a three splits, perfectly balanced class-wise: train: 49,000 out of the original 50,000 samples from the training set of CIFAR10; calibration: 1,000 left-out samples from the training set; test: 10,000 samples, the entire original test set. File Structure Files are archives <split>/<classname>.zip. Each… See the full description on the dataset page: https://huggingface.co/datasets/ego-thales/cifar10.imageimage-classification100K<n<1M0 likes338 downloads1y agoHugging Face02orion-ai-lab /Thalia Thalia: A Global, Multi-Modal Dataset for Volcanic Activity Monitoring Paper | GitHub | Interactive Demo (Colab) Thalia is a global, multi-modal dataset for volcanic activity monitoring through Satellite-based Interferometric Synthetic Aperture Radar (InSAR) imagery. Building upon the Hephaestus dataset, Thalia provides higher-resolution, multi-source, and multi-temporal data in a machine-learning-ready format. Dataset Overview Thalia consists of 38 spatiotemporal… See the full description on the dataset page: https://huggingface.co/datasets/orion-ai-lab/Thalia.textimage-classification10K<n<100K3 likes336 downloads5mo agoHugging Face03thaotien /movies_CLIP_ViT-L14 🎬 Movie Frame & Caption Dataset 📖 Introduction This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted. This dataset can be used for tasks such as: Video understanding Multimodal learning (image + text) Image captioning Vision-language retrieval 📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.imageimage-classification10K<n<100K0 likes52 downloads1y agoHugging Face04Porameht /fashion-dataset-thai fashion-dataset-thai Thai-localized fashion product dataset: 44,072 product images with metadata fields translated to Thai (gender, category, sub-category, article type, base colour, season, usage). Based on the Fashion Product Images dataset (Kaggle). Format Field Description id Product id year Year productDisplayName Product name image Product image gender_th, masterCategory_th, subCategory_th, articleType_th, baseColour_th, season_th… See the full description on the dataset page: https://huggingface.co/datasets/Porameht/fashion-dataset-thai.imageimage-classification10K<n<100K0 likes24 downloads24d agoHugging Face05ThangCao /DualStream-Foundational-Manifests Dual-Stream DeepFake Foundational Baseline Connectors This repository provides standardized data connectors, download manifests, and partition splits for the 8 foundational baseline datasets used in the Dual-Stream Deepfake Detection Framework. 📊 Dual-Stream Model Allocation 🖼️ Model 1: General Vision & Signal Model (>578,000 samples) NTIRE-RobustAIGenDetection (~120,000 samples): Multi-generator synthetic artifacts (MSU 2024). CIFAKE (120,000… See the full description on the dataset page: https://huggingface.co/datasets/ThangCao/DualStream-Foundational-Manifests.textimage-classificationn<1K0 likes18 downloads2d agoHugging Face06thanglexuan /MurineCyto-Detimageimage-classification1K<n<10K1 likes13 downloads1y agoHugging Face07ThatHungarian /Cifar-ExtendedThese images were generated with an image generator I trained on the CIFAR-10 datasetThere are 250k images in totalNot recommended for serious training, only experimentation Name Format: (name)_(counter)(index).png Name: airplane, automobile, bird, cat, deer, dog, frog, horse, ship, truckCounter: every digit besides the last, range from 0 to 24999Index: each name has one, they range from 0 to 9 (airplane - 0, automobile - 1, bird - 2...) Example: bird_115102.png (Name - bird, Counter -… See the full description on the dataset page: https://huggingface.co/datasets/ThatHungarian/Cifar-Extended.imageimage-classification100K<n<1M0 likes12 downloads9mo agoHugging Face08ThanThoai9x /vietnamese-food-imagesgated vietnamese-food-images Food images collected from Google Maps restaurant reviews, with rich metadata (place, location, review, dish classification, image-quality scores). Built with the pipeline in crawl_image — crawl review photos → AI food/dish classification + quality filtering → upload. Statistics Images: 34104 Unique places: 1989 Dish classes: 33 Dish distribution dish count other_dish 5738 bun_dau_mam_tom 4148 lau 2710… See the full description on the dataset page: https://huggingface.co/datasets/ThanThoai9x/vietnamese-food-images.imageimage-classification10K<n<100K0 likes3 downloads4mo agoHugging Face09BgBatman007 /Thaliagated Thalia: A Global, Multi-Modal Dataset for Volcanic Activity Monitoring Paper | GitHub | Interactive Demo (Colab) Thalia is a global, multi-modal dataset for volcanic activity monitoring through Satellite-based Interferometric Synthetic Aperture Radar (InSAR) imagery. Building upon the Hephaestus dataset, Thalia provides higher-resolution, multi-source, and multi-temporal data in a machine-learning-ready format. Dataset Overview Thalia consists of 38 spatiotemporal… See the full description on the dataset page: https://huggingface.co/datasets/BgBatman007/Thalia.textimage-classification10K<n<100K0 likes2 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.