CoolFace
Datasetpublic

FILM6912/th-en-zh-tts-200k-enhanced

TH-EN-ZH Multi-speaker TTS Dataset (200K, RE-USE Enhanced) Speech-enhanced variant of FILM6912/th-en-zh-tts-200k. Every clip has been processed through NVIDIA RE-USE (universal speech enhancement, SEMamba) at its native sample rate, then re-encoded losslessly as FLAC (PCM_16). Same schema, same row order, same 200,000 rows (th 100k / en 50k / zh 50k): Column Type Description text string Transcript (identical to the original dataset) audio Audio Enhanced audio, FLAC… See the full description on the dataset page: https://huggingface.co/datasets/FILM6912/th-en-zh-tts-200k-enhanced.

sourceHugging Facecc0-1.0updated 2d agoView on Hugging Face
0likes438downloads
settings

This repository belongs to FILM6912 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameth-en-zh-tts-200k-enhanced
visibilitypublic
licencecc0-1.0
gatedno
ownerFILM6912
Account settings