tonibirat/Sagarmatha-ASR-Nepali-Diamond-V3
Dataset Card for Sagarmatha ASR Nepali Diamond V3 Dataset Summary Sagarmatha ASR Nepali Diamond V3 is a large-scale, production-grade Automatic Speech Recognition (ASR) dataset designed for the Nepali language. The corpus contains 265.7 hours of verified, 16 kHz audio paired with strictly normalized Devanagari transcriptions. It was compiled and curated primarily for the fine-tuning of state-of-the-art multilingual acoustic models, including OpenAI's Whisper… See the full description on the dataset page: https://huggingface.co/datasets/tonibirat/Sagarmatha-ASR-Nepali-Diamond-V3.
08
No card is published for this repository, or it could not be fetched from Hugging Face right now.
