datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Vaani-Noise-Event-Dataset
Vaani Noise Event Timestamps
Dataset Summary
Vaani Noise Event Timestamps is a derived dataset from Project Vaani, a large-scale multilingual speech initiative by IISc Bangalore and ARTPARK that captures India's linguistic diversity across all districts.
This dataset provides noise event annotations with fine-grained timestamps for the subset audio recordings from the Vaani corpus. Each entry identifies background noise categories along with their precise start… See the full description on the dataset page: https://huggingface.co/datasets/ARTPARK-IISc/Vaani-Noise-Event-Dataset.sample-filesQUT-Event-VTR-Dataset
Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation
Welcome to the official QUT-Event-VTR-Dataset dataset repository attached to the paper Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation, to be presented at the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026).
File Structure
Cite us at
Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain… See the full description on the dataset page: https://huggingface.co/datasets/gokulbnr/QUT-Event-VTR-Dataset.audio-event-classification-post-public
audio-event-classification-post-public
Sound-event and acoustic-scene classification annotations: ESC-50 (environmental), UrbanSound8K, FSD50k (50k+ events), TUT-Acoustic-Scenes-2017, DCASE-2025, NonSpeech7k (vocal sounds), VocalSound (laugh/cough/sigh). Useful for training audio LLMs on the perception substrate underneath higher-level reasoning.
Audio is not bundled in this repo. See download.sh and per-dataset data/<name>.info.json for the fetch recipe; run postlink_audio.py… See the full description on the dataset page: https://huggingface.co/datasets/vhands/audio-event-classification-post-public.Vaani_Noise_Event_TimeStamp
Vaani Noise Event Timestamps
🚧 Dataset Status: Actively Being Built
Data is being uploaded in batches. Current coverage is a subset of the final planned corpus (~167 hrs train).
Star/watch this repo to be notified of updates.
Dataset Summary
Vaani Noise Event Timestamps is a derived dataset from Project Vaani, a large-scale multilingual speech initiative by IISc Bangalore and ARTPARK that captures India's linguistic diversity across all districts.
This dataset… See the full description on the dataset page: https://huggingface.co/datasets/PavanKumarJ-ARTPARK/Vaani_Noise_Event_TimeStamp.sed-conv-eventsevent-bench
Event Bench
29-turn multi-turn speech-to-speech benchmark for evaluating voice AI models as an event planning assistant.
Part of Audio Arena, a suite of 6 benchmarks spanning 221 turns across different domains. Built by Arcada Labs.
Leaderboard | GitHub | All Benchmarks
Dataset Description
The model acts as an event planning assistant managing venue bookings, catering, and guest logistics. The conversation features cascading changes — a venue switch triggers catering… See the full description on the dataset page: https://huggingface.co/datasets/arcada-labs/event-bench.sound_eventsSpeech_Concurrent_Event_Detectionevenki-speechgenerated-sound-eventseven_speech_biblicalThis dataset consists of audiofiles with a speech in Even language.
The correspondence between text and audio is in the table metadata.csv.
The data was collected from religious texts written down by Institute for Bible Translation
Һөвки Дукундукун укчэнэкэл. Institute for Bible Translation, Moscow, 2018.
Притчал. Institute for Bible Translation, Moscow, 2019.
Sourse
Dialect
Total length (min)
Religious texts
Lamunkhin
67.96
Another dataset of Even speech:
field records of… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_biblical.even_speech_hseThis dataset consists of audiofiles with a speech in Even language.
The correspondence between text and audio is in the table metadata.csv.
The data was collected during field trips of HSE University expedition "Languages and Cultures of Kamchatka"
Sourse
Dialect
Total length (min)
Expedition records
Bystraja
TBA
Another dataset of Even speech:
biblical texts: https://huggingface.co/datasets/tbkazakova/even_speech_biblical
There is also unified data from the project (Aralova… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_hse.benchmark-eventgpievenki-speech-translationeven_speech_pakendorfThis dataset consists of audiofiles with a speech in Even language.
The correspondence between text and audio is in the table metadata.csv.
The data was collected from the project (Aralova et al. 2007-2023)[1] and then brought to a unified format.
These are conversations about life, folklore and personal narratives, explanatory and procedural texts, individual words and phrases recorded in three districts: Bystraja District (Esso and Anavgaj villages), Sebyan-Kyuyol, Topolinoe.
This is field… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_pakendorf.audios_wavSDG_EventSound_Dataset
