datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Bo-voice-v1.0.0
Tibetan STT Benchmark Model Card
Bo-voice-v1.0.0 is a high-fidelity benchmark for Tibetan Speech-to-Text (STT) technology. It provides a rigorous, multi-domain evaluation set to measure Automatic Speech Recognition (ASR) performance across diverse acoustic environments and speaking styles.
### Dataset Overview
Snapshot Date: 15 July 2024, 02:47:06 PM
Total Samples: 8,367 audio-transcript pairs.
Verification: Every transcript has been reviewed by at least one expert in… See the full description on the dataset page: https://huggingface.co/datasets/MonlamAI/Bo-voice-v1.0.0.llama_data_v1.0.0
