CoolFace
Datasetpublic

ssz1111/SpokenWOZ-Train-Text

What is SpokenWOZ? SpokenWOZ is a large-scale multi-domain speech-text dataset for spoken task-oriented dialogue modeling, which consists of 203k turns, 5.7k dialogues and 249 hours audios from realistic human-to-human spoken conversations. Why SpokenWOZ? The majority of existing TOD datasets are constructed via writing or paraphrasing from annotators rather than being collected from realistic spoken conversations. The written TDO datasets may not be… See the full description on the dataset page: https://huggingface.co/datasets/ssz1111/SpokenWOZ-Train-Text.

sourceHugging Faceupdated 9mo agoView on Hugging Face
0likes399downloads

ssz1111/SpokenWOZ-Train-Text · main · files are served by the source, never re-hosted here