datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OpenSLR54-Nepali-ASR-parquet
OpenSLR 54: Large Nepali ASR training data set (unmodified parquet repackaging)
This is an unofficial repackaging of the official OpenSLR 54 release
(SLR54, https://www.openslr.org/54/), converted to parquet so it can be streamed with 🤗 datasets.
It is not affiliated with or endorsed by OpenSLR or the original authors.
All credit for the data belongs to the original creators (see Citation).
What's inside
157,905 utterances, 16 shards: one per original zip… See the full description on the dataset page: https://huggingface.co/datasets/JeevanDai/OpenSLR54-Nepali-ASR-parquet.OpenSLR43-Nepali-TTS-parquet
OpenSLR 43: High quality TTS data for Nepali (unmodified parquet repackaging)
This is an unofficial repackaging of the official OpenSLR 43 release (SLR43, https://www.openslr.org/43/),
converted to parquet for 🤗 datasets. It is not affiliated with or endorsed by OpenSLR, Google or the original authors.
All credit for the data belongs to the original creators (see Citation).
What's inside
2,064 utterances, 18 female speakers, ~2.8 h, 48 kHz mono; recorded in… See the full description on the dataset page: https://huggingface.co/datasets/JeevanDai/OpenSLR43-Nepali-TTS-parquet.
