datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ru-book-mix-10h
ru-book-mix-10h
A 10-hour synthetic Russian-audiobook diarization benchmark. 600 one-minute
FLAC clips (16 kHz mono, 16-bit, lossless) with NIST RTTM ground truth, generated by
mexus/diarization-benchmark
from its5Q/biggest-ru-book (speech) and bilguun/musan-noise (background).
Intended use: diarization evaluation only. This dataset is not
suitable for training — the same source voices repeat across files, so any
model that trains on it will leak voice identity into its test split.… See the full description on the dataset page: https://huggingface.co/datasets/mexus/ru-book-mix-10h.rubbish
