datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cifar10
Dataset Specifications
Contains the entire CIFAR10 dataset, downloaded via PyTorch, then split and saved as .png files representing 32x32 images.
There a three splits, perfectly balanced class-wise:
train: 49,000 out of the original 50,000 samples from the training set of CIFAR10;
calibration: 1,000 left-out samples from the training set;
test: 10,000 samples, the entire original test set.
File Structure
Files are archives <split>/<classname>.zip. Each… See the full description on the dataset page: https://huggingface.co/datasets/ego-thales/cifar10.Thalia
Thalia: A Global, Multi-Modal Dataset for Volcanic Activity Monitoring
Paper | GitHub | Interactive Demo (Colab)
Thalia is a global, multi-modal dataset for volcanic activity monitoring through Satellite-based Interferometric Synthetic Aperture Radar (InSAR) imagery. Building upon the Hephaestus dataset, Thalia provides higher-resolution, multi-source, and multi-temporal data in a machine-learning-ready format.
Dataset Overview
Thalia consists of 38 spatiotemporal… See the full description on the dataset page: https://huggingface.co/datasets/orion-ai-lab/Thalia.movies_CLIP_ViT-L14
🎬 Movie Frame & Caption Dataset
📖 Introduction
This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted.
This dataset can be used for tasks such as:
Video understanding
Multimodal learning (image + text)
Image captioning
Vision-language retrieval
📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.fashion-dataset-thai
fashion-dataset-thai
Thai-localized fashion product dataset: 44,072 product images with metadata fields translated to Thai (gender, category, sub-category, article type, base colour, season, usage). Based on the Fashion Product Images dataset (Kaggle).
Format
Field
Description
id
Product id
year
Year
productDisplayName
Product name
image
Product image
gender_th, masterCategory_th, subCategory_th, articleType_th, baseColour_th, season_th… See the full description on the dataset page: https://huggingface.co/datasets/Porameht/fashion-dataset-thai.DualStream-Foundational-Manifests
Dual-Stream DeepFake Foundational Baseline Connectors
This repository provides standardized data connectors, download manifests, and partition splits for the 8 foundational baseline datasets used in the Dual-Stream Deepfake Detection Framework.
📊 Dual-Stream Model Allocation
🖼️ Model 1: General Vision & Signal Model (>578,000 samples)
NTIRE-RobustAIGenDetection (~120,000 samples): Multi-generator synthetic artifacts (MSU 2024).
CIFAKE (120,000… See the full description on the dataset page: https://huggingface.co/datasets/ThangCao/DualStream-Foundational-Manifests.MurineCyto-DetCifar-ExtendedThese images were generated with an image generator I trained on the CIFAR-10 datasetThere are 250k images in totalNot recommended for serious training, only experimentation
Name Format:
(name)_(counter)(index).png
Name: airplane, automobile, bird, cat, deer, dog, frog, horse, ship, truckCounter: every digit besides the last, range from 0 to 24999Index: each name has one, they range from 0 to 9 (airplane - 0, automobile - 1, bird - 2...)
Example: bird_115102.png (Name - bird, Counter -… See the full description on the dataset page: https://huggingface.co/datasets/ThatHungarian/Cifar-Extended.vietnamese-food-images
vietnamese-food-images
Food images collected from Google Maps restaurant reviews, with rich metadata
(place, location, review, dish classification, image-quality scores).
Built with the pipeline in
crawl_image — crawl review photos → AI food/dish
classification + quality filtering → upload.
Statistics
Images: 34104
Unique places: 1989
Dish classes: 33
Dish distribution
dish
count
other_dish
5738
bun_dau_mam_tom
4148
lau
2710… See the full description on the dataset page: https://huggingface.co/datasets/ThanThoai9x/vietnamese-food-images.Thalia
Thalia: A Global, Multi-Modal Dataset for Volcanic Activity Monitoring
Paper | GitHub | Interactive Demo (Colab)
Thalia is a global, multi-modal dataset for volcanic activity monitoring through Satellite-based Interferometric Synthetic Aperture Radar (InSAR) imagery. Building upon the Hephaestus dataset, Thalia provides higher-resolution, multi-source, and multi-temporal data in a machine-learning-ready format.
Dataset Overview
Thalia consists of 38 spatiotemporal… See the full description on the dataset page: https://huggingface.co/datasets/BgBatman007/Thalia.
