CoolFace
Datasetpublic

yilele/synth-qa-taste-codec-chat

Synthetic QA Taste-S Codec Chat 18571 single-turn Traditional Chinese QA utterances with synthesized speech, 21.6 hours of audio before codec extraction. Assistant speech is represented as: <SAY> text_token <a_code> <b_code> ... <p_code> ... </SAY> Each text token is followed by its 16 Taste-S FSQ codes (codebooks a..p). Configurations default — messages (user question + assistant <SAY> speech), audio, and answer text. Statistics Utterances: 18571… See the full description on the dataset page: https://huggingface.co/datasets/yilele/synth-qa-taste-codec-chat.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes15downloads
6 commits on main
ae47c0e2mo ago

Upload README.md with huggingface_hub

yilele
4aa6dd12mo ago

Upload README.md with huggingface_hub

yilele
98bdfbb2mo ago

Upload dataset

yilele
554cf142mo ago

Upload README.md with huggingface_hub

yilele
8eda9022mo ago

Upload dataset

yilele
3dbb8f62mo ago

initial commit

yilele