CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alefiury /Echoes-Platos-CaveEchoes in Plato's Cave — Controlled Speech–Text Corpus Controlled corpus of 14,400 synthetic English utterances in which the same 600 sentences are rendered by 6 speakers × 4 emotions, so that speaker identity and prosody vary while linguistic content is held fixed. It was built for the paper: Echoes in Plato's Cave: Measuring Global and Local Alignment Between Speech and Language Representations, accepted as an oral presentation at the Speech and Audio Language… See the full description on the dataset page: https://huggingface.co/datasets/alefiury/Echoes-Platos-Cave.audiotext-to-speech10K<n<100K0 likes238 downloads9d agoHugging Face02alexsdl /EchoLensgated EchoLens A Human-Speech Dataset for Auditing Demographic Sensitivity in Audio-Language Models 📄 Paper (EMNLP 2026 Findings) · 💻 Code Voice interfaces are increasingly moving away from transcription pipelines toward end-to-end systems that directly respond to audio inputs. This development in turn requires a shift in evaluation methodology away from transcription accuracy and towards more substantive markers such as response validity. We introduce EchoLens, a demographically… See the full description on the dataset page: https://huggingface.co/datasets/alexsdl/EchoLens.audioaudio-text-to-text100K<n<1M0 likes180 downloads10d agoHugging Face03upb-nlp /echogated Echo Echo is a speech dataset for Romanian language crowd-sourced from the community. The dataset contains over 300 hours of speech data from 300 speakers and is available for non-commercial research purposes only. The dataset is collected using the Echo platform. audioautomatic-speech-recognition100K<n<1M4 likes129 downloads2y agoHugging Face04dawahealth /Zambezi_ECHO_v1 Zambezi ECHO v1: Shona-English Code-Switched Maternal Health Queries Dataset Description Zambezi ECHO v1 SESB (Shona-English Speech Benchmark) is a dataset of short, simulated patient voice queries in Shona (Zimbabwe), code-switched with English, covering common maternal and child health concerns — pregnancy symptoms, danger signs, child illness, and general health questions asked the way patients actually phrase them in the field, mixing Shona with English… See the full description on the dataset page: https://huggingface.co/datasets/dawahealth/Zambezi_ECHO_v1.audioautomatic-speech-recognitionn<1K0 likes69 downloads6d agoHugging Face05tarirozw /Zambezi_ECHO_v1 Zambezi ECHO v1: Shona-English Code-Switched Maternal Health Queries Dataset Description Zambezi ECHO v1 SESB (Shona-English Speech Benchmark) is a dataset of short, simulated patient voice queries in Shona (Zimbabwe), code-switched with English, covering common maternal and child health concerns — pregnancy symptoms, danger signs, child illness, and general health questions asked the way patients actually phrase them in the field, mixing Shona with English… See the full description on the dataset page: https://huggingface.co/datasets/tarirozw/Zambezi_ECHO_v1.audioautomatic-speech-recognitionn<1K0 likes60 downloads11d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.