CoolFace
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AISHELL /HI-MIAaudio0 likes1.5k downloads3y agoHugging Face02KeisukeImoto /MIAO 🐾 MIAO: Multimodal Image-Audio Onomatopoeia Dataset Dataset Summary MIAO is a multimodal dataset consisting of paired sound event clips and onomatopoeic images, which is designed to support research and development on multimodal correspondence between sounds and visual onomatopoeic expressions. It can be used for a wide range of tasks, including cross-modal retrieval, multimodal representation learning, and generative modeling of sounds and images.… See the full description on the dataset page: https://huggingface.co/datasets/KeisukeImoto/MIAO.audion<1K1 likes324 downloads4mo agoHugging Face03BrunoHays /Bangor-Miami-Spanish-English-Corpus Bangor Miami Spanish-English Corpus The Bangor Miami Corpus is a naturalistic Spanish-English code-switching speech dataset collected by Jon Russell Herring at Bangor University. It captures spontaneous bilingual conversations recorded in Miami, Florida, involving proficient Spanish-English bilinguals across multiple speaker groups. Dataset description Total recordings 56 Total duration ~32 h Languages English (en), Spanish (es) Format MP3 audio +… See the full description on the dataset page: https://huggingface.co/datasets/BrunoHays/Bangor-Miami-Spanish-English-Corpus.audioautomatic-speech-recognitionn<1K0 likes136 downloads4mo agoHugging Face04miaocongxin /KeSpeech This dataset only contains test data, which is integrated into UltraEval-Audio(https://github.com/OpenBMB/UltraEval-Audio) framework. python audio_evals/main.py --dataset KeSpeech --model gpt4o_audio 🚀超凡体验,尽在UltraEval-Audio🚀 UltraEval-Audio——全球首个同时支持语音理解和语音生成评估的开源框架,专为语音大模型评估打造,集合了34项权威Benchmark,覆盖语音、声音、医疗及音乐四大领域,支持十种语言,涵盖十二类任务。选择UltraEval-Audio,您将体验到前所未有的便捷与高效: 一键式基准管理 📥:告别繁琐的手动下载与数据处理,UltraEval-Audio为您自动化完成这一切,轻松获取所需基准测试数据。 内置评估利器… See the full description on the dataset page: https://huggingface.co/datasets/miaocongxin/KeSpeech.audio10K<n<100K0 likes127 downloads8mo agoHugging Face05potsawee /audio-mia-batch-20260312 Audio MIA Batch 20260312 This dataset contains 6,998 audio files (64 GB) downloaded from YouTube videos. Dataset Structure Each row contains: audio: Audio bytes (playable in the dataset viewer) video_id: YouTube video ID category: Content category source_term: Search term used query: Full search query title: Video title url: YouTube URL uploader: Channel name channel_id: YouTube channel ID upload_date: Upload date (YYYY-MM-DD) duration: Video duration in seconds… See the full description on the dataset page: https://huggingface.co/datasets/potsawee/audio-mia-batch-20260312.audioaudio-classification1K<n<10K0 likes76 downloads7mo agoHugging Face06mteb /MIAO-I2Aaudio10K<n<100K0 likes61 downloads2mo agoHugging Face07mteb /MIAO-A2Iaudio10K<n<100K0 likes52 downloads2mo agoHugging Face08Wissam42 /MIAO-A2Iaudio10K<n<100K0 likes37 downloads2mo agoHugging Face09minhthien /mia-meeting MIA Meeting E2E Dataset Synthetic meeting dataset for end-to-end experiments: audio to transcript transcript plus roster to action items action item extraction benchmark Splits train: 200 samples, 0 with linked audio validation: 5 samples, 5 with linked audio eval: 205 samples, 5 with linked audio Structure data/*.jsonl # split manifests audio/<split>/* # linked audio files when available transcripts/<split>/*.json #… See the full description on the dataset page: https://huggingface.co/datasets/minhthien/mia-meeting.audioautomatic-speech-recognition0 likes32 downloads4mo agoHugging Face10drewoodward /miami-corpus Bangor Miami Corpus (merged) This repository contains a merged version of the Bangor Miami Corpus of Spanish–English bilingual speech: miamiCorpus_merged.mp3 — all 56 audio recordings concatenated into a single ~35-hour MP3 file. miamiCorpus_merged.vtt — the corresponding transcripts concatenated into a single WebVTT file. This is not the canonical version. The canonical corpus (56 separate .wav / .cha file pairs in CHAT format, with gloss and translation tiers) is available from… See the full description on the dataset page: https://huggingface.co/datasets/drewoodward/miami-corpus.audioautomatic-speech-recognitionn<1K0 likes27 downloads5mo agoHugging Face11Wissam42 /MIAO-I2Aaudio10K<n<100K0 likes25 downloads2mo agoHugging Face12Luisr-ecu /miami_corpus Bangor Miami Corpus (merged) This repository contains a merged version of the Bangor Miami Corpus of Spanish–English bilingual speech: miamiCorpus_merged.mp3 — all 56 audio recordings concatenated into a single ~35-hour MP3 file. miamiCorpus_merged.vtt — the corresponding transcripts concatenated into a single WebVTT file. This is not the canonical version. The canonical corpus (56 separate .wav / .cha file pairs in CHAT format, with gloss and translation tiers) is available… See the full description on the dataset page: https://huggingface.co/datasets/Luisr-ecu/miami_corpus.audioautomatic-speech-recognitionn<1K0 likes21 downloads1mo agoHugging Face13mia-project /child_handpicked_sentencesaudion<1K0 likes16 downloads3y agoHugging Face14ZANIT /MiaNFSMWaudion<1K0 likes6 downloads3y agoHugging Face15Littleharmony /Miaaudion<1K0 likes6 downloads2y agoHugging Face16miasimsdois /badgyalaudion<1K0 likes6 downloads1y agoHugging Face17archivartaunik /uladzimir-karatkevich-byli-u-miane-miadzvedzi Былі ў мяне мядзведзі... Metadata Author: Уладзімір Караткевіч Title: Былі ў мяне мядзведзі... Narrator: Source Group: Дзіцячыя Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/uladzimir-karatkevich-byli-u-miane-miadzvedzi.audion<1K0 likes6 downloads4mo agoHugging Face18archivartaunik /ianka-sipakou-leta-z-miatlushkai-aleg-sidorchyk Лета з мятлушкай Metadata Author: Янка Сіпакоў Title: Лета з мятлушкай Narrator: Алег Сідорчык Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ianka-sipakou-leta-z-miatlushkai-aleg-sidorchyk.audion<1K0 likes6 downloads4mo agoHugging Face19Vinnyyw /Miacolucciaudion<1K0 likes5 downloads3y agoHugging Face20Berlinda /Miaaudion<1K0 likes4 downloads3y agoHugging Face21Jesynelson /Miaaudion<1K0 likes4 downloads3y agoHugging Face22archivartaunik /alena-masla-miane-zavuts-lakhneska-viktar-manaeu Мяне завуць Лахнэска Metadata Author: Алена Масла Title: Мяне завуць Лахнэска Narrator: Віктар Манаеў Source Group: Дзіцячыя Source: https://knizhnyvoz.by/ Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/alena-masla-miane-zavuts-lakhneska-viktar-manaeu.audion<1K0 likes4 downloads4mo agoHugging Face23miasimsdois /jayvaqueraudion<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.