CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tungnguyenlam /vietnamese-acoustic-boundary-verifier-data Vietnamese Acoustic Boundary & Speaker Purity Dataset (Gemini 3.8 Flash Distilled) This dataset contains 202 curated Vietnamese audio samples with fine-grained acoustic boundary annotations distilled from Google Gemini 3.8 Flash (thinkingLevel="LOW"). It is specifically designed to train and evaluate multimodal models (e.g., Gemma 4 E4B Audio) on acoustic quality control for speech synthesis and speaker diarization pipelines. Dataset Structure Each sample is… See the full description on the dataset page: https://huggingface.co/datasets/tungnguyenlam/vietnamese-acoustic-boundary-verifier-data.audioaudio-classificationn<1K0 likes136 downloads18d agoHugging Face02yangwang825 /vox1-veri-full VoxCeleb 1 VoxCeleb1 contains over 100,000 utterances for 1,251 celebrities, extracted from videos uploaded to YouTube. Verification Split train validation test # of speakers 1211 1211 40 # of samples 133777 14865 4874 References https://www.robots.ox.ac.uk/~vgg/data/voxceleb/vox1.html tabularaudio-classification100K<n<1M1 likes66 downloads3y agoHugging Face03KYAGABA /kinyarwanda_cleaned_testset_verified_20HRSaudio10K<n<100K0 likes62 downloads2y agoHugging Face04KYAGABA /kinyarwanda_cleaned_testset_verified_200HRSaudio100K<n<1M0 likes40 downloads2y agoHugging Face05MUGEN-Benchmark /Speaker_Verificationaudion<1K0 likes37 downloads8mo agoHugging Face06yangwang825 /vox2-veri-full VoxCeleb 2 VoxCeleb2 contains over 1 million utterances for 6,112 celebrities, extracted from videos uploaded to YouTube. Verification Split train validation test # of speakers 5,994 5,994 118 # of samples 982,808 109,201 36,237 Data Fields ID (string): The ID of the sample with format <spk_id--utt_id_start_stop>. duration (float64): The duration of the segment in seconds. wav (string): The filepath of the waveform. start (int64): The… See the full description on the dataset page: https://huggingface.co/datasets/yangwang825/vox2-veri-full.tabularaudio-classification1M<n<10M0 likes30 downloads3y agoHugging Face07yangwang825 /vox2-veri-3s VoxCeleb 2 VoxCeleb2 contains over 1 million utterances for 6,112 celebrities, extracted from videos uploaded to YouTube. Verification Split train validation test # of speakers 5,994 5,994 118 # of samples 982,808 109,201 36,237 Data Fields ID (string): The ID of the sample with format <spk_id--utt_id_start_stop>. duration (float64): The duration of the segment in seconds. wav (string): The filepath of the waveform. start (int64): The… See the full description on the dataset page: https://huggingface.co/datasets/yangwang825/vox2-veri-3s.tabularaudio-classification1M<n<10M0 likes25 downloads3y agoHugging Face08KYAGABA /amharic_cleaned_testset_verifiedaudio10K<n<100K1 likes25 downloads2y agoHugging Face09binbin123 /thuyg20-tts-clean-verifiedaudio1K<n<10K0 likes25 downloads1y agoHugging Face10KYAGABA /kinyarwanda_cleaned_testset_verifiedaudio100K<n<1M0 likes21 downloads2y agoHugging Face11mehmedadymn /havacilik-veriseti ATC Veri Kümesi - Whisper Modeli ile İnce Ayar Bu veri kümesi, OpenAI'nin Whisper modelini, Hava Trafik Kontrolü (ATC) iletişimlerinde transkripsiyon doğruluğunu artırmak amacıyla ince ayar yapmak için oluşturulmuştur. Veri kümesi, iki ana kaynaktan alınan transkripsiyonlar ve karşılık gelen ses dosyalarını içermektedir: ATCO2 ve UWB-ATCC korpusu, özellikle havacılıkla ilgili iletişimler için seçilmiştir. Veri kümesi, Otomatik Konuşma Tanıma (ASR) projelerinde kullanılmak üzere… See the full description on the dataset page: https://huggingface.co/datasets/mehmedadymn/havacilik-veriseti.audioautomatic-speech-recognition10K<n<100K0 likes21 downloads2y agoHugging Face12anivenu /gaia-verifiedgated GAIA-Verified A corrected subset of GAIA's validation split: 147 of the original 165 tasks. 27 of GAIA's 165 validation tasks (16.4%) are defective. 18 could not be salvaged and were removed; 9 had a gold answer that is simply wrong and were corrected. Questions were never rewritten and the scorer was never patched. Defective 27 / 165 (16.4%) Removed 18 Golds corrected 9 GAIA-Verified 147 tasks Every verdict, with its evidence, is in the audit:… See the full description on the dataset page: https://huggingface.co/datasets/anivenu/gaia-verified.audioquestion-answeringn<1K0 likes21 downloads2mo agoHugging Face13KYAGABA /kinyarwanda_cleaned_testset_verified_100HRSaudio10K<n<100K0 likes19 downloads2y agoHugging Face14Thanarit /Thai-Voice-Test-Verification Thanarit/Thai-Voice Combined Thai audio dataset from multiple sources Dataset Details Total samples: 10 Total duration: 0.01 hours Language: Thai (th) Audio format: 16kHz mono WAV Volume normalization: -20dB Sources Processed 1 datasets in streaming mode Source Datasets GigaSpeech2: Large-scale multilingual speech corpus Usage from datasets import load_dataset # Load with streaming to avoid downloading everything dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Thanarit/Thai-Voice-Test-Verification.audion<1K0 likes15 downloads1y agoHugging Face15yangwang825 /vox1-veri-3s VoxCeleb 1 VoxCeleb1 contains over 100,000 utterances for 1,251 celebrities, extracted from videos uploaded to YouTube. Verification Split train validation test # of speakers 1211 1211 40 # of samples 299246 33672 4874 References https://www.robots.ox.ac.uk/~vgg/data/voxceleb/vox1.html tabularaudio-classification100K<n<1M0 likes13 downloads3y agoHugging Face16kstunlp /kyrgyz-asr-verified-v1gated🇰🇬 Kyrgyz ASR-Verified Speech Corpus metric value Clips 411,834 Audio 874 hours Speakers 280 Verification CER mean 0.79%, median 0.00%, p90 2.35% License & access KSTU members only. This dataset is released under a custom KSTU Internal Dataset License (other, see the LICENSE file): 🎓 Access and use are restricted to KSTU (Kyrgyz State Technical University) members and KSTU-affiliated persons. 🚫 Requests from outside KSTU will not be approved. Request… See the full description on the dataset page: https://huggingface.co/datasets/kstunlp/kyrgyz-asr-verified-v1.audioautomatic-speech-recognition100K<n<1M0 likes13 downloads3mo agoHugging Face17vericudebuget /livestreamaudion<1K0 likes11 downloads1y agoHugging Face18KYAGABA /kinyarwanda_cleaned_testset_verified_10HRSaudio1K<n<10K0 likes10 downloads2y agoHugging Face19MehmetAliRenda /ses_verisiaudion<1K0 likes10 downloads8mo agoHugging Face20dmusingu /wolof-kallaama-external-verificationaudion<1K0 likes7 downloads2y agoHugging Face21abhiram4572 /VeriSpeak VeriSpeak VeriSpeak is a spoken-statement factual-verification benchmark. Each example is a short synthesized speech clip of a single declarative sentence about a public figure, labeled correct or incorrect depending on whether the spoken statement is factually true. The task: given the audio (and optionally its transcript), decide whether the claim it makes is accurate. It targets speech-native fact-checking / hallucination detection. Dataset at a glance… See the full description on the dataset page: https://huggingface.co/datasets/abhiram4572/VeriSpeak.audioaudio-classification1K<n<10K1 likes7 downloads26d agoHugging Face22KaniTTS-research-team /speaker-overlap-verifiedaudion<1K0 likes5 downloads6mo agoHugging Face23quinnlue /sonyc_ust_verifiedaudio1K<n<10K0 likes5 downloads5mo agoHugging Face24dmusingu /wolof-google-fleurs-external-verificationaudion<1K0 likes4 downloads2y agoHugging Face25tavantai /speaker_verification_dataset_collectionaudio10K<n<100K0 likes4 downloads2mo agoHugging Face26BDanial /iaaa-speaker-verificationgatedaudio1K<n<10K0 likes3 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.