CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mythicinfinity /librispeech-pc-44khz-opus LibriSpeech-PC 44kHz Opus Summary This dataset is a high-quality audio replacement variant of Librispeech PC. It preserves the row identity and text fields while replacing audio content from the source audio with the highest available quality (usually mp3 128kpbs) which is then encoded as Opus (64 kbps). Sampling rate is increased from 16khz up to 48khz depending the on source audio. LibriSpeech-PC is a merge of openslr/librispeech_asr audio metadata with SLR145… See the full description on the dataset page: https://huggingface.co/datasets/mythicinfinity/librispeech-pc-44khz-opus.audioautomatic-speech-recognition100K<n<1M5 likes413 downloads6mo agoHugging Face02speech-uk /voa-opus Voice of America for 🇺🇦 Ukrainian (OPUS) Community Discord: https://bit.ly/discord-uds Speech Recognition: https://t.me/speech_recognition_uk Speech Synthesis: https://t.me/speech_synthesis_uk Stats Total files processed: 326174 Total duration: 390h 59m 54s Other Labels generated by https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3 audioautomatic-speech-recognition100K<n<1M0 likes365 downloads6mo agoHugging Face03speech-uk /voa-2-opus Voice of America 2 for 🇺🇦 Ukrainian (OPUS) Community Discord: https://bit.ly/discord-uds Speech Recognition: https://t.me/speech_recognition_uk Speech Synthesis: https://t.me/speech_synthesis_uk Stats Total files processed: x Total duration: x audioautomatic-speech-recognition100K<n<1M0 likes359 downloads6mo agoHugging Face04speech-uk /cv22-opus Common Voice for 🇺🇦 Ukrainian (OPUS) Ukrainian validated subset of Common Voice 22 Community Discord: https://bit.ly/discord-uds Speech Recognition: https://t.me/speech_recognition_uk Speech Synthesis: https://t.me/speech_synthesis_uk Stats Total files processed: 89248 Total duration: 115h 5m 9s audioautomatic-speech-recognition10K<n<100K0 likes222 downloads6mo agoHugging Face05speech-uk /yodas2-opus YODAS2 for 🇺🇦 Ukrainian (OPUS) Ukrainian validated subset of YODAS2 Community Discord: https://bit.ly/discord-uds Speech Recognition: https://t.me/speech_recognition_uk Speech Synthesis: https://t.me/speech_synthesis_uk Stats Total files processed: 400213 Total duration: 998h 41m 3s textautomatic-speech-recognition100K<n<1M0 likes198 downloads6mo agoHugging Face06speech-uk /broadcast-opus Broadcast for 🇺🇦 Ukrainian (in OPUS) Community Discord: https://bit.ly/discord-uds Speech Recognition: https://t.me/speech_recognition_uk Speech Synthesis: https://t.me/speech_synthesis_uk Stats Total files processed: 136736 Total duration: 300h 10m 51s Other Labels generated by https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3 audioautomatic-speech-recognition100K<n<1M0 likes187 downloads6mo agoHugging Face07ivrit-ai /audio-v2-opusgatedThis dataset contains >20k hours of Hebrew audio, all licensed under the ivrit.ai v1 license. It was released on April 20th, 2025. You can find the full list of sources in this dataset under the dataset's sources.txt. Paper: https://arxiv.org/abs/2307.08720 If you use our datasets, the following quote is preferable: @misc{marmor2023ivritai, title={ivrit.ai: A Comprehensive Dataset of Hebrew Speech for AI Research and Development}, author={Yanir Marmor and Kinneret Misgav and Yair… See the full description on the dataset page: https://huggingface.co/datasets/ivrit-ai/audio-v2-opus.audioaudio-classification10K<n<100K0 likes60 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.