JacobLinCool/zh-tw-tts-comparison
zh-TW TTS comparison — audio & metadata Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl). Single source of truth for the blind-test Space and the project's GitHub Pages site. <model>/<condition>/<id>.wav — audio clips (16/24/48 kHz depending on model) sentences.jsonl — the 25 quick-test sentences (text, bucket, entities) clips.jsonl — per-clip metadata:… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/zh-tw-tts-comparison.
zh-TW TTS comparison — audio & metadata
Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl). Single source of truth for the blind-test Space and the project's GitHub Pages site.
<model>/<condition>/<id>.wav— audio clips (16/24/48 kHz depending on model)sentences.jsonl— the 25 quick-test sentences (text, bucket, entities)clips.jsonl— per-clip metadata: dual-ASR MER + hypotheses, synth time, RTF, peak VRAM
