CoolFace
Datasetpublic

Aananda-giri/openSLR-Nepali

OpenSLR Nepali Speech Dataset (Preprocessed) Dataset Description This is a preprocessed version of the Nepali speech dataset from OpenSLR, ready for training speech models including Automatic Speech Recognition (ASR) and Text-to-Speech (TTS). Dataset Statistics Total Audio Files: 118,231 Total Duration: 57.34 hours Sample Rate: 16kHz Channels: Mono Format: WAV Preprocessing Applied Text Preprocessing: Text cleaning and… See the full description on the dataset page: https://huggingface.co/datasets/Aananda-giri/openSLR-Nepali.

sourceHugging Facecc-by-sa-4.0updated 10mo agoView on Hugging Face
0likes12downloads
4 commits on main
eb4bcb610mo ago

Upload dataset files from openSLR-Nepali-augmented

Aananda-giri
d32d53310mo ago

uploading README.md and metadata.csv

Aananda-giri
afb4d3410mo ago

Upload nepali-dataset.zip with huggingface_hub

Aananda-giri
a704e6a10mo ago

initial commit

Aananda-giri