CoolFace
20 results

csr

BantuLanguagesInitiative /CSRC Congolese Speech Radio Corpus (CSRC) The Congolese Speech Radio Corpus (CSRC) is an unlabelled radio-speech corpus covering Lingala, Kikongo, and Tshiluba. This Hugging Face release was prepared by Bantu Languages Initiative from the CSRC component of Speech Recognition Datasets for Congolese Languages. Its purpose is to make the radio archives easier to use for self-supervised speech learning, ASR pretraining, acoustic adaptation, language identification, and robust speech… See the full description on the dataset page: https://huggingface.co/datasets/BantuLanguagesInitiative/CSRC.audioautomatic-speech-recognition10K<n<100K1 likes580 downloads3mo agoHugging Facearchitojha /m3-retrieve-csrimage100K<n<1M0 likes379 downloads4mo agoHugging FaceGEM /cs_restaurantsThe task is generating responses in the context of a (hypothetical) dialogue system that provides information about restaurants. The input is a basic intent/dialogue act type and a list of slots (attributes) and their values. The output is a natural language sentence.text1K<n<10K1 likes356 downloads4y agoHugging Facemoyix /debian_csrctext1M<n<10M1 likes332 downloads4y agoHugging Faceladybugdb /ldbc-csr1B<n<10B0 likes194 downloads1mo agoHugging Facecommunity-datasets /cs_restaurants Dataset Card for Czech Restaurant Dataset Summary This is a dataset for NLG in task-oriented spoken dialogue systems with Czech as the target language. It originated as a translation of the English San Francisco Restaurants dataset by Wen et al. (2015). The domain is restaurant information in Prague, with random/fictional values. It includes input dialogue acts and the corresponding outputs in Czech. Supported Tasks and Leaderboards other-intent-to-text:… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/cs_restaurants.texttext-generation1K<n<10K2 likes186 downloads2y agoHugging Face