CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01speechbrain /LoquaciousSet LargeScaleASR: 25,000 hours of transcribed and heterogeneous English speech recognition data for research and commercial use. The full details are available in the paper. Made of 6 subsets: large contains 25,000 hours of read / spontaneous and clean / noisy transcribed speech. medium contains 2,500 hours of read / spontaneous and clean / noisy transcribed speech. small contains 250 hours of read / spontaneous and clean / noisy transcribed speech. clean contains 13,000 hours of read… See the full description on the dataset page: https://huggingface.co/datasets/speechbrain/LoquaciousSet.audioautomatic-speech-recognition10M<n<100M63 likes6.8k downloads7mo agoHugging Face02sdelangen /speechbrain-samplesaudio0 likes67 downloads2y agoHugging Face03sujalappa /speech-brain-noise-evaluation-dataset Speech Brain Noise Evaluation Dataset Dataset Description This dataset contains 2,000 samples organized across multiple splits and 20 subsets. The dataset includes audio data. Dataset Structure Subsets This dataset includes the following subsets: noisy-bg-snr-10: 100 samples test: 100 samples noisy-bg-snr-30: 100 samples test: 100 samples noisy-bg-snr-50: 100 samples test: 100 samples denoised-bg-snr-10: 100 samples test: 100 samples… See the full description on the dataset page: https://huggingface.co/datasets/sujalappa/speech-brain-noise-evaluation-dataset.audioautomatic-speech-recognition1K<n<10K0 likes35 downloads1y agoHugging Face04nguyenduykha /Noises-Dataset-SpeechBrainaudion<1K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.