datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nepali-oov-distilled
Nepali OOV-distilled subset (854 h)
An OOV-dense distillation of Premal-12/c9nepali-audio-dataset2 (used with the
author's permission), shipped in four variants: the original single-voice audio,
a CPU-augmented copy, and 244 h re-rendered onto 1,842 real human speakers with
Seed-VC. For Nepali ASR and TTS work.
Filter with the variant field -- see Composition below. If you came here
for speaker diversity, you want variant == "vc".
What this is
The source corpus is… See the full description on the dataset page: https://huggingface.co/datasets/milanakdj/nepali-oov-distilled.nepali-asr-benchmark
Nepali ASR Benchmark
Per-utterance reference, hypothesis, WER, and CER for the six released Nepali ASR
checkpoints evaluated on three independent test sets. Released alongside the paper
Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech
Recognition.
Contents
Field
Type
Description
utterance_id
string
stable identifier {test_set}-{index}
reference
string
NFC-normalised gold transcription (Devanagari)
hypothesis
string… See the full description on the dataset page: https://huggingface.co/datasets/sumanpaudel1997/nepali-asr-benchmark.
