CoolFace
Datasetpublicgated

tonibirat/Sagarmatha-ASR-Nepali-Diamond-V3

Dataset Card for Sagarmatha ASR Nepali Diamond V3 Dataset Summary Sagarmatha ASR Nepali Diamond V3 is a large-scale, production-grade Automatic Speech Recognition (ASR) dataset designed for the Nepali language. The corpus contains 265.7 hours of verified, 16 kHz audio paired with strictly normalized Devanagari transcriptions. It was compiled and curated primarily for the fine-tuning of state-of-the-art multilingual acoustic models, including OpenAI's Whisper… See the full description on the dataset page: https://huggingface.co/datasets/tonibirat/Sagarmatha-ASR-Nepali-Diamond-V3.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
0likes8downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.