CoolFace
Datasetpublic

BoburAmirov/podcasts_tashkent_dialect_youtube_uzbek_speech_dataset

Tashkent dialect focused podcasts youtube uzbek speech Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with mostly tashkent dialects. The data was collected from publicly available podcast videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Jahongir Latipov interviews and Bu podcast (respect authors) YouTube… See the full description on the dataset page: https://huggingface.co/datasets/BoburAmirov/podcasts_tashkent_dialect_youtube_uzbek_speech_dataset.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
0likes13downloads
settings

This repository belongs to BoburAmirov on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namepodcasts_tashkent_dialect_youtube_uzbek_speech_dataset
visibilitypublic
licenceapache-2.0
gatedno
ownerBoburAmirov
Account settings
BoburAmirov/podcasts_tashkent_dialect_youtube_uzbek_speech_dataset · CoolFace