CoolFace
Datasetpublic

FILM6912/th-en-zh-tts-200k

TH-EN-ZH Multi-speaker TTS Dataset (200K) A unified multi-speaker text-to-speech dataset combining three high-quality speech corpora, normalized to a single schema: Column Type Description text string Transcript (whitespace-normalized; LibriTTS uses normalized text) audio Audio Embedded audio bytes (original format: MP3 for th / WAV for en, zh) speaker_id string Namespaced speaker ID (cv_th_*, libritts_*, aishell3_*) lang string th, en, or zh… See the full description on the dataset page: https://huggingface.co/datasets/FILM6912/th-en-zh-tts-200k.

sourceHugging Facecc0-1.0updated 8d agoView on Hugging Face
0likes104downloads

FILM6912/th-en-zh-tts-200k · main · files are served by the source, never re-hosted here