CoolFace
Datasetpublic

BAAI/CS-Dialogue

CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition Introduction CS-Dialogue is a large-scale, publicly available Mandarin-English code-switching speech dialogue dataset. This dataset solves key problems found in existing code-switching speech datasets — mainly their small size, lack of natural conversations, and missing full-length dialogue recordings. It provides a solid foundation for advancing… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/CS-Dialogue.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
10likes1.6kdownloads
settings

This repository belongs to BAAI on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameCS-Dialogue
visibilitypublic
licencecc-by-nc-sa-4.0
gatedno
ownerBAAI
Account settings