CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01FluidInference /THCHS-30-tests THCHS-30 Test Set THCHS-30 test split for Mandarin Chinese speech recognition benchmarking. Dataset Info Language: Mandarin Chinese (zh-CN) Samples: 2,495 Speakers: 10 Sample Rate: 16 kHz License: Apache 2.0 Usage from datasets import load_dataset # After uploading to HuggingFace dataset = load_dataset("your-username/thchs30-test") # Example print(dataset['train'][0]) # { # 'audio': {'array': [...], 'sampling_rate': 16000, 'path': 'audio/D11_750.wav'}, #… See the full description on the dataset page: https://huggingface.co/datasets/FluidInference/THCHS-30-tests.audio1K<n<10K0 likes608 downloads6mo agoHugging Face02Eureka-Leo /MCABSA_testsetaudion<1K2 likes324 downloads1y agoHugging Face03VibroNav /2024.09.14_AGH_manual_Temperature_tests_20-55_Degrees Temperature tests (20-55 Degrees) Dataset author: Oğuzhan Berke Özdil What was measured? Methods Type of needle: Quincke 22G 9cm Needle Was attached with 3d printed needle attachment directly in front of the sound hole of the microphone Type of microphone: How the measurements were taken: manual Phantoms thickness of foams : 3 cm Foam is the same as we used in Magdeburg foam phantom. Foams were soaked in warm/hot water to create different… See the full description on the dataset page: https://huggingface.co/datasets/VibroNav/2024.09.14_AGH_manual_Temperature_tests_20-55_Degrees.audion<1K0 likes207 downloads9mo agoHugging Face04CocoBro /MMEdit-TestSet MMEdit Test Set A paired audio editing test set for text-guided audio manipulation evaluation, released with MMEdit. Overview This dataset contains 3,317 aligned triplets: Component Description raw/ Source audio before editing target/ Target audio after editing content.jsonl Editing instruction (caption) keyed by audio_id Each sample is linked by a shared audio_id. For example, sample add_017221 corresponds to: raw/add_017221.wav — original… See the full description on the dataset page: https://huggingface.co/datasets/CocoBro/MMEdit-TestSet.audioaudio-to-audio1K<n<10K0 likes196 downloads4mo agoHugging Face05hoangducanh1865 /vi-en-ast-testsetaudio1K<n<10K0 likes105 downloads20d agoHugging Face06hoangducanh1865 /en-vi-ast-testsetaudio1K<n<10K0 likes69 downloads16d agoHugging Face07XRXRX /X-Voice-TestsetX-Voice Multilingual Test Set High-Fidelity Test Set for Multilingual Text-to-Speech across 30 Languages This test set is built as part of the research: X-Voice: One Speaker, 30+ Languages with Zero-Shot Voice Cloning, serving as the evaluation benchmark for our model. Dataset Summary 30 languages European: bg (Bulgarian), cs (Czech), da (Danish), de (German), el (Greek), en (English), es (Spanish), et (Estonian), fi (Finnish), fr (French), hr (Croatian), hu (Hungarian), it… See the full description on the dataset page: https://huggingface.co/datasets/XRXRX/X-Voice-Testset.audiotext-to-speech10K<n<100K4 likes65 downloads5mo agoHugging Face08KYAGABA /kinyarwanda_cleaned_testset_verified_20HRSaudio10K<n<100K0 likes49 downloads2y agoHugging Face09Yehor /cv10-uk-testset-clean The cleaned Common Voice 10 (test set) that has been checked by a human for Ukrainian 🇺🇦 Overview This repository contains the archive of Common Voice 10 (test set) with checked Ukrainian transcriptions and audios. All audios have been checked by a human to be sure that they are correct. This archive is used to test all ASR models listed here: https://github.com/egorsmkv/speech-recognition-uk Community Discord: https://bit.ly/discord-uds Speech… See the full description on the dataset page: https://huggingface.co/datasets/Yehor/cv10-uk-testset-clean.audioautomatic-speech-recognition1K<n<10K3 likes43 downloads2y agoHugging Face10KYAGABA /kinyarwanda_cleaned_testset_verified_200HRSaudio100K<n<1M0 likes40 downloads2y agoHugging Face11EYEDOL /swahili_small_testSwahilidata_77audio1K<n<10K0 likes37 downloads1y agoHugging Face12EYEDOL /swahili_small_testSwahilidata_88audio1K<n<10K0 likes31 downloads1y agoHugging Face13asoria /test_stats_erroraudio10K<n<100K0 likes30 downloads2y agoHugging Face14prvInSpace /eval_framework_testsetaudio1K<n<10K0 likes28 downloads1y agoHugging Face15EYEDOL /swahili_small_testSwahilidata_22audio1K<n<10K0 likes28 downloads1y agoHugging Face16EYEDOL /swahili_small_testSwahilidata_66audio1K<n<10K0 likes28 downloads1y agoHugging Face17Mazino0 /test_speechaudion<1K0 likes26 downloads1y agoHugging Face18EYEDOL /swahili_small_testSwahilidata_55audio1K<n<10K0 likes26 downloads1y agoHugging Face19Sammau /test_setaudio1K<n<10K0 likes25 downloads1y agoHugging Face20EYEDOL /swahili_small_testSwahilidata_11audio1K<n<10K0 likes25 downloads1y agoHugging Face21EYEDOL /swahili_small_testSwahilidata_33audio1K<n<10K0 likes25 downloads1y agoHugging Face22NagaSaiAbhinay /whisperkit_testsAll files are from: earnings22 Rencoded to 24kbps MP3 using: ffmpeg -i 4446796.wav -vn -map_metadata -1 -ac 1 -c:a libmp3lame -b:a 24k -application voip -y 4446796.mp3 audion<1K0 likes24 downloads2y agoHugging Face23Roy229 /hf7192-ocr-testsuite-504791audion<1K0 likes24 downloads1mo agoHugging Face24EYEDOL /swahili_small_testSwahilidata_44audio1K<n<10K0 likes23 downloads1y agoHugging Face25npat1509 /TestSynthaudion<1K1 likes21 downloads2mo agoHugging Face26KYAGABA /kinyarwanda_cleaned_testset_verifiedaudio100K<n<1M0 likes20 downloads2y agoHugging Face27jykim310 /audio_testsetaudion<1K0 likes19 downloads2y agoHugging Face28KYAGABA /amharic_cleaned_testset_verifiedaudio10K<n<100K1 likes17 downloads2y agoHugging Face29KYAGABA /kinyarwanda_cleaned_testset_verified_100HRSaudio10K<n<100K0 likes17 downloads2y agoHugging Face30KYAGABA /amharic_cleaned_testset_fleurs_currentaudio1K<n<10K0 likes17 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.