CoolFace
Datasetpublic

Harland/ADQA-Bench

ADQA-Bench: Audio-Dependent Question Answering Evaluation Benchmark This is the official Evaluation Set for DCASE 2026 Challenge Task 5: Audio-Dependent Question Answering (ADQA). The ADQA task focuses on addressing "Textual Hallucination" in Large Audio-Language Models (LALMs) — where models pass audio understanding benchmarks by relying on text prompts and internal linguistic priors rather than actual audio perception. ADQA introduces a rigorous evaluation framework… See the full description on the dataset page: https://huggingface.co/datasets/Harland/ADQA-Bench.

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
2likes958downloads
4 commits on main
963aa1c2mo ago

Replace eval-noanswer.jsonl with eval.jsonl (with answers); update README

Harland
d2c27342mo ago

Update README.md

Harland
68e09db4mo ago

Update README: remove answer-related content for competition integrity

Harland
6086e6a4mo ago

Upload eval_audios and update README for ADQA-Bench rename

Harland