datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
uae_collectedUAT_Dataua-polit-tinyua-speechUAE_100KUATuaspeech_tts_all_severityuaspeech_tts_very_lowuae_chatterbox_training_v0uaspeechUAE18000UAE17000listen2scene-stuttgart-outdooruaspeech_train_casted2025-zwesui-g02-medyczna
ZWESUI 2025 - Grupa 2 - medyczna (mowa syntetyczna)
Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu
Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2025 (pierwsza), tryb niestacjonarny.
Zespol (atrybucja): Grupa 2 (2025)
Zrodlo oryginalne: https://huggingface.co/datasets/yanvoi/med_male_female_r2
Domena: medyczna
Opis: 200 zdań z terminologią medyczną, mowa syntetyczna (ElevenLabs), głosy męskie i żeńskie; zbiór użyty… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2025-zwesui-g02-medyczna.UAE_validation_WAVua-speech-sr16000alsallom_update_para_UAE_transcription_low_chunk_by_elevenlab2026-dwesui-g04-admedvoice
DWESUI 2026 - Grupa 4 - ADMEDVOICE (medyczna)
Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu
Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2026, tryb dzienny.
Zespol (atrybucja): Grupa 4 (DWESUI 2026)
Zrodlo oryginalne: https://huggingface.co/datasets/TCA/zwesui-grupa-4-medyczn
Domena: medyczna
Licencja zrodla: nagrania YouTube CC-BY + Kaggle ADMEDVOICE + TTS
Status: kopia publiczna w organizacji kursowej (zespół… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2026-dwesui-g04-admedvoice.audiobooks_ua_test
About dataset
It is a dataset of ukrainian audiobooksEach sample contain an approximately 8 seconds od ukrainian speech
Loading script
>>> load_dataset("Zarakun/audiobooks_ua_test")
Dataset structure
**Every example has the following:
audio - the waveformrate - the sampling rate of the waveformfile_id - the id of the speakerduration - the duration of the video in secondssentence - the transcript of the video
uaspeech_female
Uaspeech Female Dataset
Overview
This dataset contains dysarthric speech samples from a female speaker (F02) in the UA-Speech corpus, prepared for pathological speech synthesis research.
Speaker Information:
Speaker ID: F02
Corpus: UA-Speech
Gender: Female
Speech Status: Dysarthric
Dataset Statistics
Total Samples: 1,200
Total Duration: 1.59 hours
Sampling Rate: 24,000 Hz
Format: Audio arrays with transcriptions
Training Split
Samples: 1,000… See the full description on the dataset page: https://huggingface.co/datasets/resproj007/uaspeech_female.uaspeech_tts_lowUA-SER
UA-SER: Ukrainian Speech Emotion Recognition Corpus
A labelled Ukrainian emotional speech corpus of 952 clips across four emotion classes, collected and annotated for the purpose of training and evaluating Speech Emotion Recognition (SER) models on Ukrainian.
Dataset Summary
Ukrainian is a low-resource language with no publicly available emotional speech dataset. UA-SER fills this gap by providing short naturalistic utterances labelled by three native Ukrainian annotators… See the full description on the dataset page: https://huggingface.co/datasets/OlhaHavryliuk/UA-SER.al_salloom_UAE_transcription_by_elevenlab2025-zwesui-g03-debata-prezydencka
ZWESUI 2025 - Grupa 3 - debata prezydencka 2025
Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu
Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2025 (pierwsza), tryb niestacjonarny.
Zespol (atrybucja): Grupa 3 (2025)
Zrodlo oryginalne: https://huggingface.co/datasets/directtt/polish_presidential_debate
Domena: debata prezydencka
Opis: Debata prezydencka TVP z 12 maja 2025 — 13 kandydatów, 195 wypowiedzi, mowa spontaniczna.… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2025-zwesui-g03-debata-prezydencka.uaspeech_test_casteduaspeech_male
Uaspeech Male Dataset
Overview
This dataset contains dysarthric speech samples from a male speaker (M04) in the UA-Speech corpus, prepared for pathological speech synthesis research.
Speaker Information:
Speaker ID: M04
Corpus: UA-Speech
Gender: Male
Speech Status: Dysarthric
Dataset Statistics
Total Samples: 1,200
Total Duration: 0.84 hours
Sampling Rate: 24,000 Hz
Format: Audio arrays with transcriptions
Training Split
Samples: 1,000
Duration:… See the full description on the dataset page: https://huggingface.co/datasets/resproj007/uaspeech_male.2026-dwesui-g01-neurologia
DWESUI 2026 - Grupa 1 - neurologia
Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu
Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2026, tryb dzienny.
Zespol (atrybucja): Grupa 1 (DWESUI 2026)
Zrodlo oryginalne: https://huggingface.co/datasets/JankesTNJ/dwesui-grupa-1-neurologia
Domena: neurologia
Licencja zrodla: nagrania YouTube CC-BY + synteza TTS
Status: kopia publiczna w organizacji kursowej (zespół opublikował zbiór… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2026-dwesui-g01-neurologia.UAE_validation_mp3al_sallom_UAE_transcription_by_elevenlab_tashkeel
