datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
librispeech-pc-44khz-opus
LibriSpeech-PC 44kHz Opus
Summary
This dataset is a high-quality audio replacement variant of Librispeech PC. It preserves the row identity and text fields while replacing audio content from the source audio with the highest available quality (usually mp3 128kpbs) which is then encoded as Opus (64 kbps). Sampling rate is increased from 16khz up to 48khz depending the on source audio.
LibriSpeech-PC is a merge of openslr/librispeech_asr audio metadata with SLR145… See the full description on the dataset page: https://huggingface.co/datasets/mythicinfinity/librispeech-pc-44khz-opus.voa-opus
Voice of America for 🇺🇦 Ukrainian (OPUS)
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
Stats
Total files processed: 326174
Total duration: 390h 59m 54s
Other
Labels generated by https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3
voa-2-opus
Voice of America 2 for 🇺🇦 Ukrainian (OPUS)
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
Stats
Total files processed: x
Total duration: x
cv22-opus
Common Voice for 🇺🇦 Ukrainian (OPUS)
Ukrainian validated subset of Common Voice 22
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
Stats
Total files processed: 89248
Total duration: 115h 5m 9s
yodas2-opus
YODAS2 for 🇺🇦 Ukrainian (OPUS)
Ukrainian validated subset of YODAS2
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
Stats
Total files processed: 400213
Total duration: 998h 41m 3s
broadcast-opus
Broadcast for 🇺🇦 Ukrainian (in OPUS)
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
Stats
Total files processed: 136736
Total duration: 300h 10m 51s
Other
Labels generated by https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3
audio-v2-opusThis dataset contains >20k hours of Hebrew audio, all licensed under the ivrit.ai v1 license.
It was released on April 20th, 2025.
You can find the full list of sources in this dataset under the dataset's sources.txt.
Paper: https://arxiv.org/abs/2307.08720
If you use our datasets, the following quote is preferable:
@misc{marmor2023ivritai,
title={ivrit.ai: A Comprehensive Dataset of Hebrew Speech for AI Research and Development},
author={Yanir Marmor and Kinneret Misgav and Yair… See the full description on the dataset page: https://huggingface.co/datasets/ivrit-ai/audio-v2-opus.
