speechbrain
LoquaciousSet
LargeScaleASR: 25,000 hours of transcribed and heterogeneous English speech recognition data for research and commercial use.
The full details are available in the paper.
Made of 6 subsets:
large contains 25,000 hours of read / spontaneous and clean / noisy transcribed speech.
medium contains 2,500 hours of read / spontaneous and clean / noisy transcribed speech.
small contains 250 hours of read / spontaneous and clean / noisy transcribed speech.
clean contains 13,000 hours of read… See the full description on the dataset page: https://huggingface.co/datasets/speechbrain/LoquaciousSet.common_languageThis dataset is composed of speech recordings from languages that were carefully selected from the CommonVoice database.
The total duration of audio recordings is 45.1 hours (i.e., 1 hour of material for each language).
The dataset has been extracted from CommonVoice to train language-id systems.speechbrain-samplesvi-xvector-speechbrainspeech-brain-noise-evaluation-dataset
Speech Brain Noise Evaluation Dataset
Dataset Description
This dataset contains 2,000 samples organized across multiple splits and 20 subsets.
The dataset includes audio data.
Dataset Structure
Subsets
This dataset includes the following subsets:
noisy-bg-snr-10: 100 samples
test: 100 samples
noisy-bg-snr-30: 100 samples
test: 100 samples
noisy-bg-snr-50: 100 samples
test: 100 samples
denoised-bg-snr-10: 100 samples
test: 100 samples… See the full description on the dataset page: https://huggingface.co/datasets/sujalappa/speech-brain-noise-evaluation-dataset.Noises-Dataset-SpeechBrain
