CoolFace
19 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ARTPARK-IISc /Vaani-Noise-Event-Datasetgated Vaani Noise Event Timestamps Dataset Summary Vaani Noise Event Timestamps is a derived dataset from Project Vaani, a large-scale multilingual speech initiative by IISc Bangalore and ARTPARK that captures India's linguistic diversity across all districts. This dataset provides noise event annotations with fine-grained timestamps for the subset audio recordings from the Vaani corpus. Each entry identifies background noise categories along with their precise start… See the full description on the dataset page: https://huggingface.co/datasets/ARTPARK-IISc/Vaani-Noise-Event-Dataset.audioaudio-classification10K<n<100K17 likes4.8k downloads2mo agoHugging Face02Eventual-Inc /sample-filesaudion<1K0 likes1.4k downloads1y agoHugging Face03gokulbnr /QUT-Event-VTR-Dataset Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation Welcome to the official QUT-Event-VTR-Dataset dataset repository attached to the paper Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation, to be presented at the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026). File Structure Cite us at Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain… See the full description on the dataset page: https://huggingface.co/datasets/gokulbnr/QUT-Event-VTR-Dataset.audion<1K0 likes339 downloads3mo agoHugging Face04vhands /audio-event-classification-post-public audio-event-classification-post-public Sound-event and acoustic-scene classification annotations: ESC-50 (environmental), UrbanSound8K, FSD50k (50k+ events), TUT-Acoustic-Scenes-2017, DCASE-2025, NonSpeech7k (vocal sounds), VocalSound (laugh/cough/sigh). Useful for training audio LLMs on the perception substrate underneath higher-level reasoning. Audio is not bundled in this repo. See download.sh and per-dataset data/<name>.info.json for the fetch recipe; run postlink_audio.py… See the full description on the dataset page: https://huggingface.co/datasets/vhands/audio-event-classification-post-public.textaudio-classification100K<n<1M1 likes193 downloads3mo agoHugging Face05PavanKumarJ-ARTPARK /Vaani_Noise_Event_TimeStamp Vaani Noise Event Timestamps 🚧 Dataset Status: Actively Being Built Data is being uploaded in batches. Current coverage is a subset of the final planned corpus (~167 hrs train). Star/watch this repo to be notified of updates. Dataset Summary Vaani Noise Event Timestamps is a derived dataset from Project Vaani, a large-scale multilingual speech initiative by IISc Bangalore and ARTPARK that captures India's linguistic diversity across all districts. This dataset… See the full description on the dataset page: https://huggingface.co/datasets/PavanKumarJ-ARTPARK/Vaani_Noise_Event_TimeStamp.audioaudio-classificationn<1K2 likes59 downloads4mo agoHugging Face06gijs /sed-conv-eventsaudio100K<n<1M0 likes42 downloads11mo agoHugging Face07arcada-labs /event-bench Event Bench 29-turn multi-turn speech-to-speech benchmark for evaluating voice AI models as an event planning assistant. Part of Audio Arena, a suite of 6 benchmarks spanning 221 turns across different domains. Built by Arcada Labs. Leaderboard | GitHub | All Benchmarks Dataset Description The model acts as an event planning assistant managing venue bookings, catering, and guest logistics. The conversation features cascading changes — a venue switch triggers catering… See the full description on the dataset page: https://huggingface.co/datasets/arcada-labs/event-bench.audioautomatic-speech-recognitionn<1K2 likes42 downloads6mo agoHugging Face08sdialog /sound_eventsaudion<1K0 likes35 downloads2mo agoHugging Face09MUGEN-Benchmark /Speech_Concurrent_Event_Detectionaudion<1K0 likes34 downloads8mo agoHugging Face10siberian-lang-lab /evenki-speechaudio1K<n<10K0 likes26 downloads1y agoHugging Face11laion /generated-sound-eventsaudio100K<n<1M1 likes25 downloads11mo agoHugging Face12tbkazakova /even_speech_biblicalThis dataset consists of audiofiles with a speech in Even language. The correspondence between text and audio is in the table metadata.csv. The data was collected from religious texts written down by Institute for Bible Translation Һөвки Дукундукун укчэнэкэл. Institute for Bible Translation, Moscow, 2018. Притчал. Institute for Bible Translation, Moscow, 2019. Sourse Dialect Total length (min) Religious texts Lamunkhin 67.96 Another dataset of Even speech: field records of… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_biblical.audioautomatic-speech-recognition1K<n<10K0 likes20 downloads2y agoHugging Face13tbkazakova /even_speech_hseThis dataset consists of audiofiles with a speech in Even language. The correspondence between text and audio is in the table metadata.csv. The data was collected during field trips of HSE University expedition "Languages and Cultures of Kamchatka" Sourse Dialect Total length (min) Expedition records Bystraja TBA Another dataset of Even speech: biblical texts: https://huggingface.co/datasets/tbkazakova/even_speech_biblical There is also unified data from the project (Aralova… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_hse.audioautomatic-speech-recognitionn<1K0 likes11 downloads2y agoHugging Face14SPARCO-project /benchmark-eventaudion<1K0 likes11 downloads5mo agoHugging Face15Evening2k /gpiaudion<1K0 likes8 downloads3y agoHugging Face16siberian-lang-lab /evenki-speech-translationaudio1K<n<10K0 likes8 downloads3mo agoHugging Face17tbkazakova /even_speech_pakendorfgatedThis dataset consists of audiofiles with a speech in Even language. The correspondence between text and audio is in the table metadata.csv. The data was collected from the project (Aralova et al. 2007-2023)[1] and then brought to a unified format. These are conversations about life, folklore and personal narratives, explanatory and procedural texts, individual words and phrases recorded in three districts: Bystraja District (Esso and Anavgaj villages), Sebyan-Kyuyol, Topolinoe. This is field… See the full description on the dataset page: https://huggingface.co/datasets/tbkazakova/even_speech_pakendorf.audioautomatic-speech-recognition1K<n<10K0 likes5 downloads2y agoHugging Face18Evening2k /audios_wavaudion<1K0 likes4 downloads3y agoHugging Face19twuser /SDG_EventSound_Datasetaudion<1K0 likes2 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.