dcase
Datasets
All datasets matching “dcase”dcase2025-audio-qa
DCASE 2025 HuggingFace Dataset
This script creates a HuggingFace dataset from the DCASE 2025 Audio Question Answering data.
Dataset Structure
The dataset contains the following columns:
audio: Audio file (automatically converted to mono 16bit 48kHz)
question: The formatted question with choices (if applicable)
question_text: The original question text without choices
answer: The correct answer
id: Unique identifier for each example
audio_url: Original audio URL from the… See the full description on the dataset page: https://huggingface.co/datasets/gijs/dcase2025-audio-qa.DCASE2026-Task5-DevSet
DCASE 2026 Task 5 Audio-Dependent Question Answering (ADQA) Development Set
This is the official Development Set for DCASE 2026 Challenge Task 5: Audio-Dependent Question Answering (ADQA).
The ADQA task focuses on addressing "Textual Hallucination" in Large Audio-Language Models (LALMs) — where models pass audio understanding benchmarks by relying on text prompts and internal linguistic priors rather than actual audio perception. ADQA introduces a rigorous evaluation… See the full description on the dataset page: https://huggingface.co/datasets/Harland/DCASE2026-Task5-DevSet.dcase23-task2-enriched
Dataset Card for the Enriched "DCASE 2023 Challenge Task 2 Dataset".
Dataset Summary
Data-centric AI principles have become increasingly important for real-world use cases. At Renumics we believe that classical benchmark datasets and competitions should be extended to reflect this development.
This is why we are publishing benchmark datasets with application-specific enrichments (e.g. embeddings, baseline results, uncertainties, label error scores). We hope this helps… See the full description on the dataset page: https://huggingface.co/datasets/renumics/dcase23-task2-enriched.stgram-mfn-dcase2020-dev
DCASE 2020 Task 2 — Development Dataset (STgram-MFN redistribution)
Redistribution of the DCASE 2020 Challenge Task 2 Development Dataset
(Zenodo record 3678171) in the exact
on-disk layout expected by the STgram-MFN reference runs. It contains MIMII and
ToyADMOS normal/anomalous machine sounds for unsupervised anomalous-sound
detection (ASD).
Original authors: Yuma Koizumi, Yohei Kawaguchi, Keisuke Imoto.
License: CC BY-NC-SA 4.0
(non-commercial) — inherited from the source;… See the full description on the dataset page: https://huggingface.co/datasets/LakoreAI/stgram-mfn-dcase2020-dev.2025_DCASE_AudioQA_Official
Audio SFT / Post-Training Data
The proposed audio question answering (AQA) dataset
with three categories: Bioacoustics QA (BQA), Temporal Soundscapes QA (TSQA), and Complex QA (CQA)
DCASE 2025 Task Description
Audio QA Model Baseline
Watkins Marine Mammal Sound Database
"Watkins Marine Mammal Sound Database, Woods Hole Oceanographic Institution and the New Bedford Whaling Museum."
📢 Post-Challenge Research Note
While the DCASE 2025 Challenge… See the full description on the dataset page: https://huggingface.co/datasets/PeacefulData/2025_DCASE_AudioQA_Official.stgram-mfn-dcase2020-eval
DCASE 2020 Task 2 — Evaluation + Additional Training (STgram-MFN redistribution)
Redistribution of the DCASE 2020 Challenge Task 2 Additional Training Dataset
(Zenodo 3727685) and Evaluation Dataset
(Zenodo 3841772), merged into the same
on-disk layout as the development set.
Original authors: Yuma Koizumi, Yohei Kawaguchi, Keisuke Imoto.
License: CC BY-NC-SA 4.0
(non-commercial) — inherited from the source; attribute the original authors
and keep any redistribution under the… See the full description on the dataset page: https://huggingface.co/datasets/LakoreAI/stgram-mfn-dcase2020-eval.
