datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ASCEND
Dataset Card for ASCEND
Dataset Summary
ASCEND (A Spontaneous Chinese-English Dataset) introduces a high-quality resource of spontaneous multi-turn conversational dialogue Chinese-English code-switching corpus collected in Hong Kong. ASCEND consists of 10.62 hours of spontaneous speech with a total of ~12.3K utterances. The corpus is split into 3 sets: training, validation, and test with a ratio of 8:1:1 while maintaining a balanced gender proportion on each set.… See the full description on the dataset page: https://huggingface.co/datasets/CAiRE/ASCEND.CAMEO
CAMEO: Collection of Multilingual Emotional Speech Corpora
Dataset Description
CAMEO is a curated collection of multilingual emotional speech datasets.
It includes 13 distinct datasets with transcriptions, encompassing a total of 41,265 audio samples.
The collection features audio in eight languages: Bengali, English, French, German, Italian, Polish, Russian, and Spanish.
Example Usage
The dataset can be loaded and processed using the datasets library:
from… See the full description on the dataset page: https://huggingface.co/datasets/amu-cai/CAMEO.nEMO
nEMO: Dataset of Emotional Speech in Polish
Dataset Description
nEMO is a simulated dataset of emotional speech in the Polish language. The corpus contains over 3 hours of samples recorded with the participation of nine actors portraying six emotional states: anger, fear, happiness, sadness, surprise, and a neutral state. The text material used was carefully selected to represent the phonetics of the Polish language. The corpus is available for free under the Creative… See the full description on the dataset page: https://huggingface.co/datasets/amu-cai/nEMO.pl-asr-bigos-v2BIGOS (Benchmark Intended Grouping of Open Speech) dataset goal is to simplify access to the openly available Polish speech corpora and
enable systematic benchmarking of open and commercial Polish ASR systems.IEMOCAP
IEMOCAP — full release, utterance level
All 10,039 segmented utterances of the USC-SAIL IEMOCAP corpus with 16 kHz audio,
transcripts, consensus emotion labels, consensus VAD ratings, per-annotator raw labels,
and annotator-agreement statistics.
This is a private mirror. IEMOCAP is distributed under the USC SAIL academic license,
which requires an individually signed agreement and does not permit redistribution.
Do not make this repository public. Anyone needing the data should… See the full description on the dataset page: https://huggingface.co/datasets/cairocode/IEMOCAP.ePark_jiu_jie_jiao_cai_nine_level_materials
FormosanBank publication status
This audio is associated with XML published in the public FormosanBank corpus and uses the same license recorded in that XML: CC BY-NC-SA 4.0. View the published XML. Publication approval is recorded on the corresponding FormosanBank Basecamp card.
FormosanBank/ePark_jiu_jie_jiao_cai_nine_level_materials
Commercial AI Use is prohibited without prior written permission. See the FormosanBank Terms of Use and AI Use Addendum.
This… See the full description on the dataset page: https://huggingface.co/datasets/FormosanBank/ePark_jiu_jie_jiao_cai_nine_level_materials.
