CoolFace
Datasetpublic

chosenek/czech-speech-combined

Czech Speech Combined Dataset Quality-filtered Czech speech dataset for TTS/ASR training. 146,153 clips across ~1,700 speakers from 5 sources. Sources Source Clips Hours Speakers Origin audiobooks 24,535 ~33h 13 Czech audiobook narrations audiobooks_new 28,131 ~39h 8+ Czech audiobook narrations yodas_czech 28,208 ~37h ~2,300 YODAS YouTube speech (quality-filtered) voxpopuli_czech 12,679 ~33h 45 VoxPopuli parliament speech commonvoice_czech 52… See the full description on the dataset page: https://huggingface.co/datasets/chosenek/czech-speech-combined.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
2likes174downloads
settings

This repository belongs to chosenek on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameczech-speech-combined
visibilitypublic
licencecc-by-4.0
gatedno
ownerchosenek
Account settings
chosenek/czech-speech-combined · CoolFace