CoolFace
Datasetpublic

OvozifyLabs/asr_evaluate_set

Speech-to-Text Evaluation Dataset Dataset Overview This dataset is designed for evaluating Uzbek speech-to-text (STT) models on real-world conversational speech data. The audio samples were collected from various open Telegram groups, capturing natural voice messages in diverse acoustic conditions and speaking styles. Key Statistics Total Samples: 745 audio files Total Duration: 1 hour 40 minutes (~100 minutes) Average Duration: ~8 seconds per… See the full description on the dataset page: https://huggingface.co/datasets/OvozifyLabs/asr_evaluate_set.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
1likes37downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
OvozifyLabs/asr_evaluate_set · CoolFace