datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
l2-arctic-release-v5.0
L2-ARCTIC v5.0
L2-ARCTIC is a non-native English speech corpus intended for research in
pronunciation assessment, accent conversion, voice conversion, and
mispronunciation detection.
This Hub dataset mirrors the v5.0 release as distributed locally:
speaker-level zip archives, the suitcase corpus archive, prompts, the
original README, and the license.
Summary
24 non-native English speakers
26,867 utterances
27.1 hours of scripted speech
Manual phone-level annotations for… See the full description on the dataset page: https://huggingface.co/datasets/chikingsley/l2-arctic-release-v5.0.l2-arctic-manual-v5.0-16k
l2-arctic-manual-v5.0-16k
This dataset is a prepared derivative of L2-ARCTIC v5.0 that keeps only
the manually annotated material and converts the audio to 16 kHz mono FLAC.
It is designed to plug into the current peacock-asr training code, which
can consume a Hugging Face dataset with audio plus phonemes.
Included splits
train: 1800 rows, 1.84 hours
validation: 899 rows, 0.94 hours
test: 900 rows, 0.88 hours
suitcase: 22 rows, 0.44 hours
The scripted subset uses the… See the full description on the dataset page: https://huggingface.co/datasets/chikingsley/l2-arctic-manual-v5.0-16k.
