CoolFace
Datasetpublic

zeroweight-ai/ZeroSpeech

ZeroSpeech A large synthetic Vietnamese speech corpus for ASR training: 9,867,987 utterances / 26,896 hours, spoken by 199,265 distinct voices, generated with ZeroTTS from web and conversational text. Every clip is 16 kHz mono FLAC, 1–30 s, paired with the exact text it was synthesized from. Fields field type description audio Audio(16 kHz) the waveform, FLAC-encoded text string the transcript — the exact string given to the TTS source string which… See the full description on the dataset page: https://huggingface.co/datasets/zeroweight-ai/ZeroSpeech.

sourceHugging Facecc-by-nc-4.0updated 26d agoView on Hugging Face
0likes2.9kdownloads
settings

This repository belongs to zeroweight-ai on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameZeroSpeech
visibilitypublic
licencecc-by-nc-4.0
gatedno
ownerzeroweight-ai
Account settings
zeroweight-ai/ZeroSpeech · CoolFace