CoolFace
Datasetpublic

OvozifyLabs/asr_evaluate_set

Speech-to-Text Evaluation Dataset Dataset Overview This dataset is designed for evaluating Uzbek speech-to-text (STT) models on real-world conversational speech data. The audio samples were collected from various open Telegram groups, capturing natural voice messages in diverse acoustic conditions and speaking styles. Key Statistics Total Samples: 745 audio files Total Duration: 1 hour 40 minutes (~100 minutes) Average Duration: ~8 seconds per… See the full description on the dataset page: https://huggingface.co/datasets/OvozifyLabs/asr_evaluate_set.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
1likes40downloads
filedata-00000-of-00002.arrow362.6 MBdownload
filedata-00001-of-00002.arrow369.2 MBdownload

OvozifyLabs/asr_evaluate_set · main · files are served by the source, never re-hosted here