csr
Datasets
All datasets matching “csr”CSRC
Congolese Speech Radio Corpus (CSRC)
The Congolese Speech Radio Corpus (CSRC) is an unlabelled radio-speech corpus covering Lingala, Kikongo, and Tshiluba.
This Hugging Face release was prepared by Bantu Languages Initiative from the CSRC component of Speech Recognition Datasets for Congolese Languages. Its purpose is to make the radio archives easier to use for self-supervised speech learning, ASR pretraining, acoustic adaptation, language identification, and robust speech… See the full description on the dataset page: https://huggingface.co/datasets/BantuLanguagesInitiative/CSRC.m3-retrieve-csrcs_restaurantsThe task is generating responses in the context of a (hypothetical) dialogue
system that provides information about restaurants. The input is a basic
intent/dialogue act type and a list of slots (attributes) and their values.
The output is a natural language sentence.debian_csrcldbc-csrcs_restaurants
Dataset Card for Czech Restaurant
Dataset Summary
This is a dataset for NLG in task-oriented spoken dialogue systems with Czech as the target language. It originated as a translation of the English San Francisco Restaurants dataset by Wen et al. (2015). The domain is restaurant information in Prague, with random/fictional values. It includes input dialogue acts and the corresponding outputs in Czech.
Supported Tasks and Leaderboards
other-intent-to-text:… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/cs_restaurants.
