CoolFace
Datasetpublic

FILM6912/th-en-zh-tts-200k-enhanced

TH-EN-ZH Multi-speaker TTS Dataset (200K, RE-USE Enhanced) Speech-enhanced variant of FILM6912/th-en-zh-tts-200k. Every clip has been processed through NVIDIA RE-USE (universal speech enhancement, SEMamba) at its native sample rate, then re-encoded losslessly as FLAC (PCM_16). Same schema, same row order, same 200,000 rows (th 100k / en 50k / zh 50k): Column Type Description text string Transcript (identical to the original dataset) audio Audio Enhanced audio, FLAC… See the full description on the dataset page: https://huggingface.co/datasets/FILM6912/th-en-zh-tts-200k-enhanced.

sourceHugging Facecc0-1.0updated 2d agoView on Hugging Face
0likes438downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
FILM6912/th-en-zh-tts-200k-enhanced · CoolFace