datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vmc2026-track1-dev
vmc2026-track1-dev
Development subset of the VMC 2026 Track 1 data.
The data is organized into two configs corresponding to two subjective evaluation paradigms: absolute rating (acr) and pairwise comparison (ccr). The sample_id values are namespaced strings such as vmc2026-track1-dev-acr_489 and vmc2026-track1-dev-ccr_7233.
acr -- Absolute Category Rating
1,008 samples. Each row pairs a sample_id with one speech audio file, its released Mean Opinion Score (MOS)… See the full description on the dataset page: https://huggingface.co/datasets/urgent-challenge/vmc2026-track1-dev.vmc2026-track1-test
vmc2026-track1-test
Test subset of the VMC 2026 Track 1 data.
The data is organized into two configs corresponding to two subjective evaluation paradigms: absolute rating (acr) and pairwise comparison (ccr). The sample_id values are namespaced strings such as vmc2026-track1-test-acr_4588 and vmc2026-track1-test-ccr_3061.
acr -- Absolute Category Rating
4,032 samples. Each row pairs a sample_id with one speech audio file, its released Mean Opinion Score (MOS)… See the full description on the dataset page: https://huggingface.co/datasets/urgent-challenge/vmc2026-track1-test.
