CoolFace
Datasetpublic

hostbot77/news_youtube_uzbek_speech_dataset

News Youtube Uzbek Speech Dataset Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with differenent dialects. The data was collected from publicly available news videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Kunuz, Qalampir YouTube channels. The data was transcribed using Gemini 2.5 Pro and was… See the full description on the dataset page: https://huggingface.co/datasets/hostbot77/news_youtube_uzbek_speech_dataset.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes109downloads
settings

This repository belongs to hostbot77 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namenews_youtube_uzbek_speech_dataset
visibilitypublic
licenceapache-2.0
gatedno
ownerhostbot77
Account settings
hostbot77/news_youtube_uzbek_speech_dataset · CoolFace