CoolFace
Datasetpublic

JacobLinCool/zh-tw-tts-comparison

zh-TW TTS comparison — audio & metadata Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl). Single source of truth for the blind-test Space and the project's GitHub Pages site. <model>/<condition>/<id>.wav — audio clips (16/24/48 kHz depending on model) sentences.jsonl — the 25 quick-test sentences (text, bucket, entities) clips.jsonl — per-clip metadata:… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/zh-tw-tts-comparison.

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
0likes66downloads
Dataset Card

zh-TW TTS comparison — audio & metadata

Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl). Single source of truth for the blind-test Space and the project's GitHub Pages site.

  • <model>/<condition>/<id>.wav — audio clips (16/24/48 kHz depending on model)
  • sentences.jsonl — the 25 quick-test sentences (text, bucket, entities)
  • clips.jsonl — per-clip metadata: dual-ASR MER + hypotheses, synth time, RTF, peak VRAM