datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Creative-Professionals-Agentic-Tasks-1M
Creative Professionals Agentic Tasks (1M)
Abstract
A massive-scale, high-fidelity synthetic task dataset comprising 1,070,917 agentic command operations across 36 creative, technical, and engineering software environments. This dataset is engineered exclusively to stress-test, evaluate, and fine-tune multimodal AI agents designed for Agent Environment operation, complex software interaction, and multi-step reasoning within deep software infrastructures.… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Creative-Professionals-Agentic-Tasks-1M.Creative-Professionals-Agentic-Tasks-1M
Creative Professionals Agentic Tasks (1M)
Abstract
A massive-scale, high-fidelity synthetic task dataset comprising 1,070,917 agentic command operations across 36 creative, technical, and engineering software environments. This dataset is engineered exclusively to stress-test, evaluate, and fine-tune multimodal AI agents designed for Agent Environment operation, complex software interaction, and multi-step reasoning within deep software infrastructures.… See the full description on the dataset page: https://huggingface.co/datasets/rAVEUK/Creative-Professionals-Agentic-Tasks-1M.audio_data_kaggle_train_taskaaudio_data_kaggle_train_taskb_audio_data_kaggle_train_taskbNADI-2025-Sub-task-3-allFor training and developing your models in the closed track, we provide the following datasets, which are publicly available on Hugging Face: The datasets represent a wide range of Arabic varieties and recording conditions, with over 85K training sentences in total. The datasets consist of dialectal, modern standard, classical, and code-switched Arabic speech and transcriptions. All except the Mixat and ArzEn subset are diacritized.
Dataset
Type
Diacritized
Train
Dev
MDASPC… See the full description on the dataset page: https://huggingface.co/datasets/MBZUAI/NADI-2025-Sub-task-3-all.audio_data_kaggle_train_ne_taskcCreative-Professionals-Agentic-Tasks-1M
Creative Professionals Agentic Tasks (1M)
Abstract
A massive-scale, high-fidelity synthetic task dataset comprising 1,070,917 agentic command operations across 36 creative, technical, and engineering software environments. This dataset is engineered exclusively to stress-test, evaluate, and fine-tune multimodal AI agents designed for Agent Environment operation, complex software interaction, and multi-step reasoning within deep software infrastructures.… See the full description on the dataset page: https://huggingface.co/datasets/kryp1234/Creative-Professionals-Agentic-Tasks-1M.nadi2026-adi20-micro-25pct-knnvc
NADI 2026 ADI20-micro — kNN-VC augmented (4 target voices)
Voice-converted copy of the 25% stratified subset (seed 42) of
UBC-NLP/NADI_2026_ADI20_micro, made with kNN-VC
following Abdullah et al. 2025.
Configs: voice_01–voice_04, 16,757 rows each, train split only.
Validation/test audio is deliberately left natural.
Target voices: 4 Arabic speakers from Common Voice (~60s each), gender-balanced,
the same set used across all dialects.
column
meaning
audio
converted… See the full description on the dataset page: https://huggingface.co/datasets/nadi-task2/nadi2026-adi20-micro-25pct-knnvc.dcase2025_task2_dev
DCASE 2025 Task 2 - Development Dataset
Attributes d1v, d2v, d3v are encoded as ClassLabels.
Usage
from datasets import load_dataset
dataset = load_dataset('HTill/dcase2025_task2_dev', trust_remote_code=True)
M3-SLU-Task2dcase2016_task2_synth
Dataset Card for "dcase2016_task2_synth"
More Information needed
audio_data_kaggle_test_taskb_M3-SLU-Task1spoken-nlp-tasks-24k
Spoken Text Benchmarks for Audio LLM Evaluation
TTS-synthesized audio versions of standard NLP text benchmarks, designed for
evaluating audio/speech LLMs on tasks where the ground-truth text is known.
These datasets were originally text-only; this resource provides spoken audio
renditions so that audio LLMs can be evaluated on the same tasks and compared
against text-only baselines.
Dataset Description
This dataset contains 24,000 WAV files: 1,000 utterances x 6 TTS… See the full description on the dataset page: https://huggingface.co/datasets/jb1999/spoken-nlp-tasks-24k.audio_data_kaggle_test_taskcasr-task-datadcase24_task10_loc1
DCASE 2024 Challenge Task 10 Development Dataset: Acoustic-based Traffic Monitoring - Location 1 subset
Citation
Bondi, L., Ghaffarzadegan, S., Damiano, S., Kumar, A., Wu, H.-H., Lin, W.-C., Das, S., Horst, H.-G., & Waterschoot, T. van . (2024). DCASE 2024 Challenge Task 10 Development Dataset: Acoustic-based Traffic Monitoring [Data set]. Zenodo. https://doi.org/10.5281/zenodo.10700792
License
Creative Commons Attribution-NonCommercial-ShareAlike 4.0… See the full description on the dataset page: https://huggingface.co/datasets/renumics/dcase24_task10_loc1.asr-taskasr-task-testaudio_data_kaggle_test_taskbtry-task-translation-african-amh-enDCase2016_Task2_Hear_2021asr-task-hindiaudio_data_kaggle_test_taskaM3-SLU-Task2-sample
🎧 M3-SLU Task 2 — Sample Dataset
🗣️ Multi-Speaker, Multi-Turn, Multi-Modal Spoken Language Understanding
🌍 Overview
The M3-SLU (Task 2 Sample) dataset is part of the M3-SLU Benchmark designed to evaluate speaker-attributed reasoning in multi-speaker, multi-turn conversations.It pairs long-form audio, transcripts, and contextual metadata, enabling fair comparison between cascade (SD + ASR + LLM) and end-to-end MLLMs.
👉 This sample includes 100 instances across 4… See the full description on the dataset page: https://huggingface.co/datasets/M3-SLU/M3-SLU-Task2-sample.voice_clone_taskM3-SLU-Task1-sample
🎧 M3-SLU Task 1 — Sample Dataset
🗣️ Multi-Speaker, Multi-Turn, Multi-Modal Spoken Language Understanding
🌍 Overview
The M3-SLU (Task 1 Sample) dataset is part of the M3-SLU Benchmark designed to evaluate speaker-attributed reasoning in multi-speaker, multi-turn conversations.It pairs long-form audio, transcripts, and contextual metadata, enabling fair comparison between cascade (SD + ASR + LLM) and end-to-end MLLMs.
👉 This sample includes 100 instances across 4… See the full description on the dataset page: https://huggingface.co/datasets/M3-SLU/M3-SLU-Task1-sample.shared_task
