CoolFace
Datasetpublic

BAAI/CS-Dialogue

CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition Introduction CS-Dialogue is a large-scale, publicly available Mandarin-English code-switching speech dialogue dataset. This dataset solves key problems found in existing code-switching speech datasets — mainly their small size, lack of natural conversations, and missing full-length dialogue recordings. It provides a solid foundation for advancing… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/CS-Dialogue.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
10likes1.6kdownloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

BAAI/CS-Dialogue · main · files are served by the source, never re-hosted here