datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
common-sense-facts-audio
Common-Sense Facts Audio Dataset
A spoken fact-completion dataset for evaluating whether Speech Language Models can retrieve common-sense and factual knowledge from speech.
Each example contains three paired versions:
prompt: an incomplete factual prompt, e.g. "the capital of France is"
fact: the correct full sentence, e.g. "the capital of France is Paris"
counterfactual: an incorrect matched sentence from the same category, e.g. "the capital of France is Rome"
The dataset… See the full description on the dataset page: https://huggingface.co/datasets/slprl/common-sense-facts-audio.huawei-common-sense
