datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
HI-MIAMIAO
🐾 MIAO: Multimodal Image-Audio Onomatopoeia Dataset
Dataset Summary
MIAO is a multimodal dataset consisting of paired sound event clips and onomatopoeic images, which is designed to support research and development on multimodal correspondence between sounds and visual onomatopoeic expressions.
It can be used for a wide range of tasks, including cross-modal retrieval, multimodal representation learning, and generative modeling of sounds and images.… See the full description on the dataset page: https://huggingface.co/datasets/KeisukeImoto/MIAO.Bangor-Miami-Spanish-English-Corpus
Bangor Miami Spanish-English Corpus
The Bangor Miami Corpus is a naturalistic Spanish-English code-switching speech dataset collected by Jon Russell Herring at Bangor University. It captures spontaneous bilingual conversations recorded in Miami, Florida, involving proficient Spanish-English bilinguals across multiple speaker groups.
Dataset description
Total recordings
56
Total duration
~32 h
Languages
English (en), Spanish (es)
Format
MP3 audio +… See the full description on the dataset page: https://huggingface.co/datasets/BrunoHays/Bangor-Miami-Spanish-English-Corpus.KeSpeech
This dataset only contains test data, which is integrated into UltraEval-Audio(https://github.com/OpenBMB/UltraEval-Audio) framework.
python audio_evals/main.py --dataset KeSpeech --model gpt4o_audio
🚀超凡体验,尽在UltraEval-Audio🚀
UltraEval-Audio——全球首个同时支持语音理解和语音生成评估的开源框架,专为语音大模型评估打造,集合了34项权威Benchmark,覆盖语音、声音、医疗及音乐四大领域,支持十种语言,涵盖十二类任务。选择UltraEval-Audio,您将体验到前所未有的便捷与高效:
一键式基准管理 📥:告别繁琐的手动下载与数据处理,UltraEval-Audio为您自动化完成这一切,轻松获取所需基准测试数据。
内置评估利器… See the full description on the dataset page: https://huggingface.co/datasets/miaocongxin/KeSpeech.audio-mia-batch-20260312
Audio MIA Batch 20260312
This dataset contains 6,998 audio files (64 GB) downloaded from YouTube videos.
Dataset Structure
Each row contains:
audio: Audio bytes (playable in the dataset viewer)
video_id: YouTube video ID
category: Content category
source_term: Search term used
query: Full search query
title: Video title
url: YouTube URL
uploader: Channel name
channel_id: YouTube channel ID
upload_date: Upload date (YYYY-MM-DD)
duration: Video duration in seconds… See the full description on the dataset page: https://huggingface.co/datasets/potsawee/audio-mia-batch-20260312.MIAO-I2AMIAO-A2IMIAO-A2Imia-meeting
MIA Meeting E2E Dataset
Synthetic meeting dataset for end-to-end experiments:
audio to transcript
transcript plus roster to action items
action item extraction benchmark
Splits
train: 200 samples, 0 with linked audio
validation: 5 samples, 5 with linked audio
eval: 205 samples, 5 with linked audio
Structure
data/*.jsonl # split manifests
audio/<split>/* # linked audio files when available
transcripts/<split>/*.json #… See the full description on the dataset page: https://huggingface.co/datasets/minhthien/mia-meeting.miami-corpus
Bangor Miami Corpus (merged)
This repository contains a merged version of the Bangor Miami Corpus of Spanish–English bilingual speech:
miamiCorpus_merged.mp3 — all 56 audio recordings concatenated into a single ~35-hour MP3 file.
miamiCorpus_merged.vtt — the corresponding transcripts concatenated into a single WebVTT file.
This is not the canonical version. The canonical corpus (56 separate .wav / .cha file pairs in CHAT format, with gloss and translation tiers) is available from… See the full description on the dataset page: https://huggingface.co/datasets/drewoodward/miami-corpus.MIAO-I2Amiami_corpus
Bangor Miami Corpus (merged)
This repository contains a merged version of the Bangor Miami Corpus of Spanish–English bilingual speech:
miamiCorpus_merged.mp3 — all 56 audio recordings concatenated into a single ~35-hour MP3 file.
miamiCorpus_merged.vtt — the corresponding transcripts concatenated into a single WebVTT file.
This is not the canonical version. The canonical corpus (56 separate .wav / .cha file pairs in CHAT format, with gloss and translation tiers) is available… See the full description on the dataset page: https://huggingface.co/datasets/Luisr-ecu/miami_corpus.child_handpicked_sentencesMiaNFSMWMiabadgyaluladzimir-karatkevich-byli-u-miane-miadzvedzi
Былі ў мяне мядзведзі...
Metadata
Author: Уладзімір Караткевіч
Title: Былі ў мяне мядзведзі...
Narrator:
Source Group: Дзіцячыя
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/uladzimir-karatkevich-byli-u-miane-miadzvedzi.ianka-sipakou-leta-z-miatlushkai-aleg-sidorchyk
Лета з мятлушкай
Metadata
Author: Янка Сіпакоў
Title: Лета з мятлушкай
Narrator: Алег Сідорчык
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ianka-sipakou-leta-z-miatlushkai-aleg-sidorchyk.MiacolucciMiaMiaalena-masla-miane-zavuts-lakhneska-viktar-manaeu
Мяне завуць Лахнэска
Metadata
Author: Алена Масла
Title: Мяне завуць Лахнэска
Narrator: Віктар Манаеў
Source Group: Дзіцячыя
Source: https://knizhnyvoz.by/
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/alena-masla-miane-zavuts-lakhneska-viktar-manaeu.jayvaquer
