CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01CAiRE /ASCEND Dataset Card for ASCEND Dataset Summary ASCEND (A Spontaneous Chinese-English Dataset) introduces a high-quality resource of spontaneous multi-turn conversational dialogue Chinese-English code-switching corpus collected in Hong Kong. ASCEND consists of 10.62 hours of spontaneous speech with a total of ~12.3K utterances. The corpus is split into 3 sets: training, validation, and test with a ratio of 8:1:1 while maintaining a balanced gender proportion on each set.… See the full description on the dataset page: https://huggingface.co/datasets/CAiRE/ASCEND.audioautomatic-speech-recognition10K<n<100K53 likes1.7k downloads2y agoHugging Face02amu-cai /CAMEO CAMEO: Collection of Multilingual Emotional Speech Corpora Dataset Description CAMEO is a curated collection of multilingual emotional speech datasets. It includes 13 distinct datasets with transcriptions, encompassing a total of 41,265 audio samples. The collection features audio in eight languages: Bengali, English, French, German, Italian, Polish, Russian, and Spanish. Example Usage The dataset can be loaded and processed using the datasets library: from… See the full description on the dataset page: https://huggingface.co/datasets/amu-cai/CAMEO.audioaudio-classification10K<n<100K16 likes363 downloads1y agoHugging Face03amu-cai /nEMO nEMO: Dataset of Emotional Speech in Polish Dataset Description nEMO is a simulated dataset of emotional speech in the Polish language. The corpus contains over 3 hours of samples recorded with the participation of nine actors portraying six emotional states: anger, fear, happiness, sadness, surprise, and a neutral state. The text material used was carefully selected to represent the phonetics of the Polish language. The corpus is available for free under the Creative… See the full description on the dataset page: https://huggingface.co/datasets/amu-cai/nEMO.audioaudio-classification1K<n<10K12 likes153 downloads2y agoHugging Face04amu-cai /pl-asr-bigos-v2gatedBIGOS (Benchmark Intended Grouping of Open Speech) dataset goal is to simplify access to the openly available Polish speech corpora and enable systematic benchmarking of open and commercial Polish ASR systems.audioautomatic-speech-recognition10K<n<100K5 likes139 downloads7mo agoHugging Face05cairocode /IEMOCAP IEMOCAP — full release, utterance level All 10,039 segmented utterances of the USC-SAIL IEMOCAP corpus with 16 kHz audio, transcripts, consensus emotion labels, consensus VAD ratings, per-annotator raw labels, and annotator-agreement statistics. This is a private mirror. IEMOCAP is distributed under the USC SAIL academic license, which requires an individually signed agreement and does not permit redistribution. Do not make this repository public. Anyone needing the data should… See the full description on the dataset page: https://huggingface.co/datasets/cairocode/IEMOCAP.audioaudio-classification10K<n<100K0 likes54 downloads2d agoHugging Face06FormosanBank /ePark_jiu_jie_jiao_cai_nine_level_materials FormosanBank publication status This audio is associated with XML published in the public FormosanBank corpus and uses the same license recorded in that XML: CC BY-NC-SA 4.0. View the published XML. Publication approval is recorded on the corresponding FormosanBank Basecamp card. FormosanBank/ePark_jiu_jie_jiao_cai_nine_level_materials Commercial AI Use is prohibited without prior written permission. See the FormosanBank Terms of Use and AI Use Addendum. This… See the full description on the dataset page: https://huggingface.co/datasets/FormosanBank/ePark_jiu_jie_jiao_cai_nine_level_materials.audioautomatic-speech-recognition100K<n<1M0 likes31 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.