CoolFace
Datasetpublicgated

lianghsun/tw-hokkien-audio-qwen3

Dataset Card for tw-hokkien-audio-qwen3 tw-hokkien-audio-qwen3 是一個台語(閩南語)之合成語音資料集,由 Qwen3-TTS-12Hz-1.7B-Base 以自然語言之「voice design」模式生成。每筆資料包含音頻、台語文本、音頻長度,並附帶描述說話者之自然語言 voice design 文字,以及情境標籤(domain / subdomain / scene / emotion / accent)。本資料集目前為 pilot 批次(8 筆),作為後續大規模擴充之前的格式與管線驗證。 Dataset Details Dataset Description 本資料集之設計目的是以文字描述(voice design)取代傳統之音色 embedding 或 x-vector,直接由自然語言指令產生具台語特定腔調、情緒與場景之合成語音。每筆生成流程為: 從 lianghsun/tw-hokkien-seed-text… See the full description on the dataset page: https://huggingface.co/datasets/lianghsun/tw-hokkien-audio-qwen3.

sourceHugging Facecc-by-4.0updated 6mo agoView on Hugging Face
3likes21downloads

No commit history came back for main. The revision may not exist, or the source declined the request.