OvozifyLabs/asr_evaluate_set
Speech-to-Text Evaluation Dataset Dataset Overview This dataset is designed for evaluating Uzbek speech-to-text (STT) models on real-world conversational speech data. The audio samples were collected from various open Telegram groups, capturing natural voice messages in diverse acoustic conditions and speaking styles. Key Statistics Total Samples: 745 audio files Total Duration: 1 hour 40 minutes (~100 minutes) Average Duration: ~8 seconds per… See the full description on the dataset page: https://huggingface.co/datasets/OvozifyLabs/asr_evaluate_set.
140
