CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01hdong51 /MultiOOD MultiOOD: Scaling Out-of-Distribution Detection for Multiple Modalities Hao Dong1   Yue Zhao2   Eleni Chatzi1   Olga Fink3 1ETH Zurich, 2University of Southern California, 3EPFL • arXiv • MultiOOD is the first-of-its-kind benchmark for Multimodal OOD Detection, characterized by diverse dataset sizes and varying modality combinations. Code https://github.com/donghao51/MultiOOD MultiOOD Benchmark MultiOOD… See the full description on the dataset page: https://huggingface.co/datasets/hdong51/MultiOOD.imagefeature-extractionn<1K0 likes1k downloads2y agoHugging Face02akuzdeuov /qwen3-tts-multilingual-emotional-speechaudio1M<n<10M0 likes748 downloads14d agoHugging Face03sheng22213 /multi_round_speech_180kaudio1M<n<10M2 likes569 downloads1y agoHugging Face04smz8599 /GUI-AIMA-multiturnimage100K<n<1M1 likes388 downloads7mo agoHugging Face05WueNLP /Synthdog-Multilingual-100 Synthdog Multilingual The Synthdog dataset created for training in Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model. Using the official Synthdog code, we created >1 million training samples for improving OCR capabilities in Large Vision-Language Models. Dataset Details We provide the images for download in two .tar.gz files. Download and extract them in folders of the same name (so cat images.tar.gz.* | tar xvzf -C images; tar xvzf… See the full description on the dataset page: https://huggingface.co/datasets/WueNLP/Synthdog-Multilingual-100.imageimage-to-text1M<n<10M4 likes307 downloads2y agoHugging Face06songlab /multiz100way-pigzSource: https://huggingface.co/datasets/songlab/multiz100wayInspired by https://huggingface.co/datasets/lpigou/89.zarr 99.zarr.tar.gz is created with pigz with option -1 for fastest decompression. Install pigz if not already installed: sudo apt install pigz Download and extract: wget https://huggingface.co/datasets/gonzalobenegas/99.zarr/resolve/main/99.zarr.tar.gz unpigz < 99.zarr.tar.gz | tar -x text1M<n<10M0 likes170 downloads2y agoHugging Face07Indulge-Bai /Multiphysics_Bench Multiphysics Bench Dataset: huggingface.co/datasets/Indulge-Bai/Multiphysics_Bench Paper: Multiphysics Bench: Benchmarking and Investigating Scientific Machine Learning for Multiphysics PDEs We propose the first general multiphysics benchmark dataset that encompasses six canonical coupled scenarios across domains such as electromagnetics, heat transfer, fluid flow, solid mechanics, pressure acoustics, and mass transport. This benchmark features the most comprehensive coupling types… See the full description on the dataset page: https://huggingface.co/datasets/Indulge-Bai/Multiphysics_Bench.image100K<n<1M4 likes153 downloads1y agoHugging Face08gdsu /sdxl_images_sb_prompts-multi_artist-seed1image10K<n<100K0 likes120 downloads2y agoHugging Face09clip-benchmark /wds_voc2007_multilabelimage1K<n<10K1 likes113 downloads4y agoHugging Face10RareConcepts /ZImage-Turbo-200k-multires-aspectbucketed ZImage-Turbo WebDataset Generated images from the ZImage-Turbo model using DiffusionDB prompts. Generation Details Hardware: 8x NVIDIA RTX 3090 GPUs Generation Time: ~2 days Estimated Cost: ~$70 (cloud compute) Dataset Statistics Total Samples: 211,081 Total Shards: 216 Samples per Shard: ~1000 Shard Naming Convention Tarballs are named: {base_resolution}-{aspect_ratio}-{shard_num:04d}-of-{total_shards:04d}.tar For example:… See the full description on the dataset page: https://huggingface.co/datasets/RareConcepts/ZImage-Turbo-200k-multires-aspectbucketed.text100K<n<1M0 likes113 downloads9mo agoHugging Face11issai /Multilingual_Speech_Dataset Multilingual Speech Dataset Paper: A Study of Multilingual End-to-End Speech Recognition for Kazakh, Russian, and English Repository: https://github.com/IS2AI/MultilingualASR Description: This repository provides the dataset used in the paper "A Study of Multilingual End-to-End Speech Recognition for Kazakh, Russian, and English". The paper focuses on training a single end-to-end (E2E) ASR model for Kazakh, Russian, and English, comparing monolingual and multilingual approaches… See the full description on the dataset page: https://huggingface.co/datasets/issai/Multilingual_Speech_Dataset.audioautomatic-speech-recognition100K<n<1M3 likes110 downloads2y agoHugging Face12cagatayn /multi_accent_speech Multi-Accent English Speech Corpus (Augmented & Speaker-Disjoint) This dataset is a curated and augmented multi-accent English speech corpus designed for speech recognition, accent classification, and representation learning.It consolidates multiple open-source accent corpora, converts all audio to a unified format, applies targeted data augmentation, and exports in a tidy, Hugging Face–ready structure. ✨ Key Features Accents covered (12 total):american_english… See the full description on the dataset page: https://huggingface.co/datasets/cagatayn/multi_accent_speech.audio100K<n<1M2 likes81 downloads1y agoHugging Face13lorenzobottelli /multi-target-spacecraft-pose-estimation Multi-target Synthetic Dataset for Spacecraft Pose Estimation Overview This dataset was developed for 6D pose estimation of unseen, non-cooperative spacecraft in proximity-operations scenarios. Most existing datasets focus on a single target, which leads models to overfit to a specific spacecraft and limits their ability to generalize to previously unseen targets. To address this limitation, the present dataset is multi-target and includes a wide variety of spacecraft… See the full description on the dataset page: https://huggingface.co/datasets/lorenzobottelli/multi-target-spacecraft-pose-estimation.image10K<n<100K0 likes55 downloads7mo agoHugging Face14collabora /multilingual-librispeech-webdatasetaudio10K<n<100K1 likes53 downloads3y agoHugging Face15Mestoukirdi /modelnet40_multi_viewimage10K<n<100K0 likes47 downloads1y agoHugging Face16JSSICE /Multi-Domain-Sentiment-DatasetUsing it for assessment. Dataset for Multi Domain (Including Kitchen, Books, DVDs, and Electronics) Multi-Domain Sentiment Dataset by John Blitzer, Mark Dredze, Fernando Pereira. Description: The Multi-Domain Sentiment Dataset contains product reviews taken from Amazon.com from 4 product types (domains): Kitchen, Books, DVDs, and Electronics. Each domain has several thousand reviews, but the exact number varies by domain. Reviews contain star ratings (1 to 5 stars) that… See the full description on the dataset page: https://huggingface.co/datasets/JSSICE/Multi-Domain-Sentiment-Dataset.textn<1K0 likes40 downloads4y agoHugging Face17guangzhaoli /multilingual-test-distill-strong-tts-20260520 Multilingual Test Distill Strong TTS 20260520 This repository contains a distributable tar-sharded version of multilingual_test_distill_strong_tts_20260520. The dataset follows the local voice_dataset/data layout after extraction: data/csvs/metadata_zh.csv data/csvs/metadata_en.csv data/csvs/metadata_ja.csv data/csvs/metadata_ko.csv data/zh/**/*.wav data/en/**/*.wav data/ja/**/*.wav data/ko/**/*.wav Metadata format: file_path|duration|dnsmos|text dnsmos is intentionally blank… See the full description on the dataset page: https://huggingface.co/datasets/guangzhaoli/multilingual-test-distill-strong-tts-20260520.audiotext-to-speech1K<n<10K0 likes38 downloads4mo agoHugging Face18ZenithVoyager /MultiEventVideotext1K<n<10K0 likes30 downloads10mo agoHugging Face19tachiwin /multilingual_ocr_erniesdkimage10K<n<100K0 likes29 downloads9mo agoHugging Face20Rashidbm /multiguard-phase2-data MultiGuard Phase 2 — Multimodal Misinformation Detection Dataset & Caches This dataset packages everything needed to reproduce Phase 2 of the MultiGuard project without redoing the heavy preprocessing. What's inside File Size Contents forensic_3class.csv 3.7 MB 18,000-sample 3-class dataset (Real / Manipulated / OOC), balanced 6,000 per class, deterministic 70/15/15 per-class split (seed 42). dct_cache.tar.gz 2.9 GB Precomputed DCT maps for all 18,000 images.… See the full description on the dataset page: https://huggingface.co/datasets/Rashidbm/multiguard-phase2-data.text10K<n<100K0 likes27 downloads5mo agoHugging Face21hellodfan /multipleimage100K<n<1M0 likes26 downloads8mo agoHugging Face22ayousanz /piper-plus-multilingual-7lang-v8-datasetgated piper-plus multilingual 7-lang v8 training dataset Preprocessed training dataset for piper-plus zero-shot TTS v8 (342,855 utterances / 3,692 speakers / 7 languages). data/ holds the full dataset (audio_norm + spec caches, split tar.gz — concatenate parts then extract). essential/ holds metadata only (dataset.jsonl + config + CAM++ speaker embeddings + holdout) for fast restore. Access is gated (manual approval) because the ja subset derives from MoeSpeech (see… See the full description on the dataset page: https://huggingface.co/datasets/ayousanz/piper-plus-multilingual-7lang-v8-dataset.text100K<n<1M0 likes24 downloads1mo agoHugging Face23suwesh /RACECAR-multislow_poliThis dataset contains the 4D point cloud data from LiDAR sensors collected from fully autonomous and self-driving Indy race cars which raced in the Indy autonomous challenge. The dataset is in nuScenes format and is divided into 7,150 sweeps and 1,199 samples which contain fused sensor data from 3 LiDARs equipped by the vehicle. This dataset's scenario is PoliMove team’s Multi-Agent Slow on LVMS racetrack. Each .pcd file contains 4 dimensional data: (x,y,z) coordinates of the 3D space and… See the full description on the dataset page: https://huggingface.co/datasets/suwesh/RACECAR-multislow_poli.text1K<n<10K0 likes23 downloads2y agoHugging Face24Evo-LMM /multimodal-open-r1-8kimage1K<n<10K0 likes21 downloads2y agoHugging Face25czyang /MultiFoley-VGGSound-Test-Audio Video-Guided Foley Sound Generation with Multimodal Controls Paper & Project page This dataset contains the generated results of our MultiFoley work on the filtered VGGSound test cases. We generate 4 samples for each 8s video (we use the first 8s video for evaluation). The results are generated with both silent video inputs and text inputs (we use the VGGSound category name for simplicity). Each wave file is named in the format of {category_name}/{u_id}_{start_time}_{idx}.wav, where… See the full description on the dataset page: https://huggingface.co/datasets/czyang/MultiFoley-VGGSound-Test-Audio.audio10K<n<100K3 likes16 downloads2y agoHugging Face26Zoooora /MultimodalSentimentOnSocialMediaimagen<1K0 likes13 downloads2y agoHugging Face27VoiceNet /multilingual-in-the-wildaudio100K<n<1M0 likes13 downloads5mo agoHugging Face28gdsu /sdxl_images_easy_prompts-multi_artist-seed0image10K<n<100K0 likes12 downloads2y agoHugging Face29lucasjin /M4-Instruct-Multiimage100K<n<1M0 likes12 downloads1y agoHugging Face30harvey59 /Multi-turn-editingimage10K<n<100K0 likes11 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.