CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01JamalLee /Omni-Fake-SET Omni-Fake-SET Omni-Fake-SET is the in-distribution split of Omni-Fake, a unified multimodal deepfake dataset for social-media forensics. It covers image, audio, video, and audio–video talking-head (AV-TH) modalities. Each modality uses the same three-way label space: real, fully synthetic, and tampered. Pair with the held-out benchmark Omni-Fake-OOD for out-of-distribution evaluation. Paper: arXiv:2605.01638 Project page: Omni-Fake License: CC-BY-4.0 Video (hybrid… See the full description on the dataset page: https://huggingface.co/datasets/JamalLee/Omni-Fake-SET.audioimage-classification1M<n<10M5 likes2.8k downloads3mo agoHugging Face02bezirganyan /LUMA LUMA A Benchmark Dataset for Learning from Uncertain and Multimodal Data 📄 📷 🎵 📊 ❓ Multimodal Uncertainty Quantification at Your Fingertips The LUMA dataset is a multimodal dataset, including audio, text, and image modalities, intended for benchmarking multimodal learning and multimodal uncertainty quantification.Paper: LUMA: A Benchmark Dataset for Learning from Uncertain and Multimodal Data Code:… See the full description on the dataset page: https://huggingface.co/datasets/bezirganyan/LUMA.audioimage-classification1K<n<10K6 likes2.6k downloads1y agoHugging Face03JamalLee /Omni-Fake-OOD Omni-Fake-OOD Omni-Fake-OOD is the out-of-distribution benchmark split of Omni-Fake. Samples come from held-out generators and platforms not included in training, for measuring cross-domain generalization. It covers image, audio, video, and audio–video talking-head (AV-TH) with the same three-class labels as Omni-Fake-SET: real, fully synthetic, and tampered. Use together with Omni-Fake-SET (in-distribution training data). Paper: arXiv:2605.01638 Project page: Omni-Fake… See the full description on the dataset page: https://huggingface.co/datasets/JamalLee/Omni-Fake-OOD.audioimage-classification100K<n<1M1 likes1.7k downloads3mo agoHugging Face04nyuuzyou /OpenGameArt-CC0 Dataset Card for OpenGameArt-CC0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons 0 (CC0) license, making them effectively public domain works. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual:… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC0.audioimage-classification10K<n<100K10 likes1.5k downloads1y agoHugging Face05Ardea /Icarus-dataset Icarus A unified multi-modal curriculum dataset for evolutionary neural architecture search. Every row is one self-contained Task = {meta, support, query}, where support and query are lists of (input_Field, output_Field) pairs. The inner loop trains on support; fitness is scored on query. Support is non-empty for every task. Encoders read the Field descriptor (axes, value_type, n_classes, value_range, mask); mask is True where a value is padding/ignored. meta.class_names, when… See the full description on the dataset page: https://huggingface.co/datasets/Ardea/Icarus-dataset.textimage-classification10K<n<100K2 likes1.2k downloads3mo agoHugging Face06vnahata /AfriMCQA-category-classification Afri-MCQA cross-modal cultural category classification (MTEB) Classify the cultural category of an entry from its photograph and the question about it spoken by a native speaker, across 16 African languages. Labels index this list: geography, building, and landmarks public figure and pop culture cooking and food objects, materials, clothing tranditions, art, and history brands, products, and companies plants and animals people, and everyday life vehicles and transportation… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/AfriMCQA-category-classification.audioaudio-classification1K<n<10K0 likes1.2k downloads19d agoHugging Face07nyuuzyou /OpenGameArt-OGA-BY-4.0 Dataset Card for OpenGameArt-OGA-BY-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the OpenGameArt Attribution 4.0 (OGA-BY-4.0) license. The dataset includes various types of game assets such as 2D art, music, sound effects, and associated metadata. Languages The dataset is primarily monolingual: English (en): All asset descriptions and metadata are in English Dataset… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-OGA-BY-4.0.audioimage-classificationn<1K2 likes687 downloads1y agoHugging Face08Hariprasath5128 /marine-animals-multimodal-dataset Marine Animals Multimodal Dataset 🐋 A comprehensive multimodal dataset combining audio recordings and images of 32 marine species. Dataset Summary Total samples: 24,911 Species: 32 Audio files: 1,357 unique recordings Images: 581 (309 matched + 272 from iNaturalist) Features species (string): Species name label (int32): Numeric label (0–31) audio (Audio): Audio recording of the species image (Image): Species image image_index (int32): Image number… See the full description on the dataset page: https://huggingface.co/datasets/Hariprasath5128/marine-animals-multimodal-dataset.audioaudio-classification10K<n<100K0 likes646 downloads9mo agoHugging Face09ccmusic-database /pianos Dataset Card for Piano Sound Quality Dataset The original dataset is sourced from the Piano Sound Quality Dataset, which includes 12 full-range audio files in .wav/.mp3/.m4a format representing seven models of pianos: Kawai upright piano, Kawai grand piano, Young Change upright piano, Hsinghai upright piano, Grand Theatre Steinway piano, Steinway grand piano, and Pearl River upright piano. Additionally, there are 1,320 split monophonic audio files in .wav/.mp3/.m4a format, bringing… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/pianos.audioaudio-classification10K<n<100K52 likes580 downloads7mo agoHugging Face10mdimamhosen /pd-voice-full-multimodal-dataset Parkinson Voice — Full Multimodal Dataset Complete Parkinson’s vs healthy voice package for classification and explainable Mel reasoning research (EDGE). Not Mel-only: raw audio, 10 visual modalities, feature CSVs, plus Gemma reasoning traces for Mel. Contents Path Description audio/ Waveform clips (healthy / parkinsons), 1134 files images/mel/ Mel spectrograms images/spectrogram/ Linear spectrograms images/mfcc/ MFCC maps images/delta_mfcc/… See the full description on the dataset page: https://huggingface.co/datasets/mdimamhosen/pd-voice-full-multimodal-dataset.imageaudio-classification10K<n<100K0 likes498 downloads14d agoHugging Face11nyuuzyou /OpenGameArt-CC-BY-SA-3.0 Dataset Card for OpenGameArt-CC-BY-SA-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution-ShareAlike 3.0 Unported (CC-BY-SA-3.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-SA-3.0.audioimage-classification1K<n<10K1 likes481 downloads1y agoHugging Face12orrzohar /EMID-Emotion-Matching EMID-Emotion-Matching orrzohar/EMID-Emotion-Matching is a derived dataset built on top of the Emotionally paired Music and Image Dataset (EMID) from ECNU (ecnu-aigc/EMID). It is designed for music ↔ image emotion matching with Qwen-Omni–style models. Each example contains: audio: mono waveform stored as datasets.Audio (HF Hub preview can play it) sampling_rate: sampling rate used when decoding (typically 16 kHz) image: a single image (datasets.Image) same: bool, whether the audio… See the full description on the dataset page: https://huggingface.co/datasets/orrzohar/EMID-Emotion-Matching.audioaudio-classification10K<n<100K0 likes433 downloads10mo agoHugging Face13nyuuzyou /OpenGameArt-OGA-BY-3.0 Dataset Card for OpenGameArt-OGA-BY-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the OpenGameArt Attribution (OGA-BY-3.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and associated metadata. Languages The dataset is primarily monolingual: English (en): All asset descriptions and metadata are in… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-OGA-BY-3.0.audioimage-classificationn<1K1 likes423 downloads1y agoHugging Face14ccmusic-database /bel_canto Dataset Card for Bel Conto and Chinese Folk Song Singing Tech Original Content This dataset is created by the authors and encompasses two distinct singing styles: bel canto and Chinese folk singing. Bel canto is a vocal technique frequently employed in Western classical music and opera, symbolizing the zenith of vocal artistry within the broader Western musical heritage. Chinese folk singing, for which there is no official English translation, is referred to here as a… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/bel_canto.audioaudio-classification10K<n<100K33 likes380 downloads7mo agoHugging Face15Skhaled /Ar-MUSA Data Directory Structure The Ar-MUSA directory contains annotated datasets organized by batches and annotation teams. Each batch is labeled with a number, and the annotation team is indicated by a letter. The structure is as follows: Ar-MUSA ├── Annotation 1a │ ├── frames # Contains the extracted frames for each record │ ├── audios # Contains the corresponding audio files │ ├── transcripts # Contains the transcripts of the audio files │ └── annotations.csv #… See the full description on the dataset page: https://huggingface.co/datasets/Skhaled/Ar-MUSA.audiotext-classification1K<n<10K3 likes369 downloads1y agoHugging Face16nyuuzyou /OpenGameArt-CC-BY-3.0 Dataset Card for OpenGameArt-CC-BY-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 3.0 (CC-BY-3.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual: English (en): All… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-3.0.textimage-classification1K<n<10K1 likes349 downloads1y agoHugging Face17anyangsong /Blockchain-Sensitive-Detect-Datagated Blockchain-Sensitive-Detect-Data English README 复旦大学附属儿科医院-区块链敏感信息检测项目的多模态完整测试数据集。 项目仓库:https://github.com/anyangsong/Blockchain-Sensitive-Detect 数据集:https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data checkpoints:https://huggingface.co/anyangsong/Blockchain-Sensitive-Detect-Checkpoints 数据以原始文件夹组织,覆盖文本、音频、图像与视频等样本。 该仓库不提供统一的 CSV、Parquet 或 JSONL 清单;类别信息主要由目录名和文件名携带。 内容警告: 数据集包含辱骂、性内容、暴力、政治相关内容、误导性医疗信息、欺诈信息。使用者应仅在具备适当访问控制、伦理审查和当地法律依据的环境中处理这些内容。… See the full description on the dataset page: https://huggingface.co/datasets/anyangsong/Blockchain-Sensitive-Detect-Data.audiotext-classificationn<1K0 likes256 downloads18d agoHugging Face18nyuuzyou /OpenGameArt-Mixed-Licenses Dataset Card for OpenGameArt-Mixed-Licenses Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are available under multiple licenses simultaneously. This dataset includes assets where creators have made their work available under two or more license options. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata, all… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-Mixed-Licenses.audioimage-classification1K<n<10K0 likes226 downloads1y agoHugging Face19blanchon /ADVANCE ADVANCE Audiovisual Aerial Scene Recognition Dataset (ADVANCE) is a comprehensive resource designed for audiovisual aerial scene recognition tasks. It consists of 5,075 pairs of geotagged audio recordings and high-resolution 512x512 RGB images extracted from FreeSound and Google Earth. These images are then labeled into 13 scene categories using OpenStreetMap. Paper: https://arxiv.org/abs/2005.08449 Homepage: https://akchen.github.io/ADVANCE-DATASET/ Description… See the full description on the dataset page: https://huggingface.co/datasets/blanchon/ADVANCE.audioimage-classification1K<n<10K1 likes154 downloads3y agoHugging Face20gfbati /Ten2ZeroThis dataset contains the following: 1- A balanced audio dataset of spoken Arabic digits from ten to zero in wav form (located at the "Dataset" folder); 2- A balanced image dataset of spoken Arabic digits from ten to zero in png form (located at the "Dataset" folder); 3- Tabular data generated using deep learning (SqueezeNet and Inception v3) from the spectrograms of the audio files; 4- Orange Data Mining workflows (".ows" files) used in processing this dataset. Please cite the following… See the full description on the dataset page: https://huggingface.co/datasets/gfbati/Ten2Zero.audioaudio-classificationn<1K1 likes131 downloads3y agoHugging Face21irfankabir02 /OpenGameArt-GPL-3.0 Dataset Card for OpenGameArt-GPL-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the GNU General Public License version 3.0 (GPL-3.0). The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, and textures along with their associated metadata. Languages The dataset is primarily monolingual: English (en): All asset descriptions… See the full description on the dataset page: https://huggingface.co/datasets/irfankabir02/OpenGameArt-GPL-3.0.audioimage-classificationn<1K0 likes127 downloads10mo agoHugging Face22nyuuzyou /OpenGameArt-CC-BY-4.0 Dataset Card for OpenGameArt-CC-BY-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 4.0 International (CC-BY-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual: English… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-4.0.audioimage-classification1K<n<10K2 likes114 downloads1y agoHugging Face23nyuuzyou /OpenGameArt-GPL-2.0 Dataset Card for OpenGameArt-GPL-2.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the GNU General Public License version 2.0 (GPL-2.0). The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily monolingual: English (en): All asset… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-GPL-2.0.audioimage-classificationn<1K0 likes96 downloads1y agoHugging Face24openlyne /OpenGameArt-CC-BY-4.0 Dataset Card for OpenGameArt-CC-BY-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution 4.0 International (CC-BY-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/openlyne/OpenGameArt-CC-BY-4.0.textimage-classification1K<n<10K0 likes71 downloads3mo agoHugging Face25deepsafe /evaluation-datasetgated DeepSafe Evaluation Dataset Evaluation set for DeepSafe, a deepfake detection benchmark. Tiers Tier Samples Generators Size Use master_eval_small/ 198 116 1.7 GB smoke test, under 2 min master_eval/ 15,454 411 10 GB the standard benchmark master_eval_full/ 45,954 411 25 GB complete set Medium tier composition: 9,954 image, 3,500 audio, 2,000 video. from huggingface_hub import snapshot_download snapshot_download("deepsafe/evaluation-dataset"… See the full description on the dataset page: https://huggingface.co/datasets/deepsafe/evaluation-dataset.audioaudio-classification10K<n<100K0 likes65 downloads3d agoHugging Face26nyuuzyou /OpenGameArt-GPL-3.0 Dataset Card for OpenGameArt-GPL-3.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the GNU General Public License version 3.0 (GPL-3.0). The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, and textures along with their associated metadata. Languages The dataset is primarily monolingual: English (en): All asset descriptions… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-GPL-3.0.audioimage-classificationn<1K0 likes60 downloads1y agoHugging Face27nyuuzyou /OpenGameArt-CC-BY-SA-4.0 Dataset Card for OpenGameArt-CC-BY-SA-4.0 Dataset Summary This dataset contains game artwork assets collected from OpenGameArt.org that are specifically released under the Creative Commons Attribution-ShareAlike 4.0 International (CC-BY-SA-4.0) license. The dataset includes various types of game assets such as 2D art, 3D art, concept art, music, sound effects, textures, and documents along with their associated metadata. Languages The dataset is primarily… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/OpenGameArt-CC-BY-SA-4.0.audioimage-classificationn<1K0 likes55 downloads1y agoHugging Face28webxos /audioform_dataset AAA UUUUUUUU UUUUUUUUDDDDDDDDDDDDD IIIIIIIIII OOOOOOOOO FFFFFFFFFFFFFFFFFFFFFF OOOOOOOOO RRRRRRRRRRRRRRRRR MMMMMMMM MMMMMMMM A:::A U::::::U U::::::UD::::::::::::DDD I::::::::I OO:::::::::OO F::::::::::::::::::::F OO:::::::::OO R::::::::::::::::R M:::::::M M:::::::M A:::::A U::::::U U::::::UD:::::::::::::::DD I::::::::I OO:::::::::::::OO… See the full description on the dataset page: https://huggingface.co/datasets/webxos/audioform_dataset.imageimage-to-textn<1K2 likes46 downloads3mo agoHugging Face2901Yassine /MNWgated MNW Benchmark Dataset Microsoft-Northwestern-WITNESS (MNW) benchmark for evaluating AI-generated media and deepfake detection across images, video, and audio. Notice This dataset is intended for evaluation purposes only. It cannot be used for training or commercial purposes. Contents Folder Description AI_Images/ AI-generated images from many generators (~15k+ samples) AI_Video/ AI-generated video samples Deepfake_Audio/ Deepfake / AI… See the full description on the dataset page: https://huggingface.co/datasets/01Yassine/MNW.audioimage-classificationn<1K0 likes37 downloads27d agoHugging Face30karankhatavkar /ak47-acoustic-rul-simulated AK-47 Acoustic Run-to-Failure (RUL) Simulation Dataset A synthetic Run-to-Failure dataset for Remaining Useful Life (RUL) estimation of an AK-47's recoil spring from gunshot audio. Because real run-to-failure recordings of a wearing firearm are practically impossible to collect, this dataset is generated by a physics-based Digital Twin that takes a small set of real, healthy gunshot recordings and mathematically simulates the acoustic signature of mechanical wear over thousands… See the full description on the dataset page: https://huggingface.co/datasets/karankhatavkar/ak47-acoustic-rul-simulated.audioaudio-classification1 likes19 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.