datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AuditoryBench
Dataset
AuditoryBench
AuditoryBench is the first dataset aimed at evaluating language models' auditory knowledge. It comprises:
Animal Sound Recognition: Predict the animal based on an onomatopoeic sound (e.g., "meow").
Sound Pitch Comparison: Compare the pitch of different sound sources.
Animal Sound Recognition
animal: The name of the animal that the sound corresponds to (e.g., cat).
description: Description of the animal sound (e.g., meow).
sentence: A sentence… See the full description on the dataset page: https://huggingface.co/datasets/HJOK/AuditoryBench.AuditoryBenchpp
AuditoryBench++
AuditoryBench++ is a benchmark designed to evaluate auditory commonsense knowledge and reasoning abilities of language models without requiring direct audio input.Humans can effortlessly reason about sounds (e.g., pitch, loudness, or animal-sound associations) even without hearing them. In contrast, language models often lack such capabilities, limiting their effectiveness in multimodal interaction.
This benchmark provides a systematic way to measure whether LLMs… See the full description on the dataset page: https://huggingface.co/datasets/HJOK/AuditoryBenchpp.
