datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Echoes-Platos-CaveEchoes in Plato's Cave — Controlled Speech–Text Corpus
Controlled corpus of 14,400 synthetic English utterances in which the same 600 sentences are
rendered by 6 speakers × 4 emotions, so that speaker identity and prosody vary
while linguistic content is held fixed. It was built for the paper: Echoes in Plato's Cave: Measuring Global and Local Alignment Between Speech and Language Representations, accepted as an oral presentation at the Speech and Audio Language… See the full description on the dataset page: https://huggingface.co/datasets/alefiury/Echoes-Platos-Cave.EchoLens
EchoLens
A Human-Speech Dataset for Auditing Demographic Sensitivity in Audio-Language Models
📄 Paper (EMNLP 2026 Findings) ·
💻 Code
Voice interfaces are increasingly moving away from transcription pipelines toward end-to-end systems that directly respond to audio inputs. This development in turn requires a shift in evaluation methodology away from transcription accuracy and towards more substantive markers such as response validity. We introduce EchoLens, a demographically… See the full description on the dataset page: https://huggingface.co/datasets/alexsdl/EchoLens.echo
Echo
Echo is a speech dataset for Romanian language crowd-sourced from the community.
The dataset contains over 300 hours of speech data from 300 speakers and is
available for non-commercial research purposes only. The dataset is collected
using the Echo platform.
Zambezi_ECHO_v1
Zambezi ECHO v1: Shona-English Code-Switched Maternal Health Queries
Dataset Description
Zambezi ECHO v1 SESB (Shona-English Speech Benchmark) is a dataset of short,
simulated patient voice queries in Shona (Zimbabwe), code-switched with
English, covering common maternal and child health concerns — pregnancy
symptoms, danger signs, child illness, and general health questions asked
the way patients actually phrase them in the field, mixing Shona with
English… See the full description on the dataset page: https://huggingface.co/datasets/dawahealth/Zambezi_ECHO_v1.Zambezi_ECHO_v1
Zambezi ECHO v1: Shona-English Code-Switched Maternal Health Queries
Dataset Description
Zambezi ECHO v1 SESB (Shona-English Speech Benchmark) is a dataset of short,
simulated patient voice queries in Shona (Zimbabwe), code-switched with
English, covering common maternal and child health concerns — pregnancy
symptoms, danger signs, child illness, and general health questions asked
the way patients actually phrase them in the field, mixing Shona with
English… See the full description on the dataset page: https://huggingface.co/datasets/tarirozw/Zambezi_ECHO_v1.
