datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
stg-paired-audioopenstbench-paired-set
OpenSTBench LibriTTS-based Paired Speaker Set
This dataset contains the LibriTTS-based paired speaker set constructed for speaker preservation evaluation in OpenSTBench, a multidimensional benchmark for speech translation systems.
It is designed to support evaluation of whether speech-to-speech translation systems preserve speaker characteristics when generating translated speech.
Paper: https://arxiv.org/abs/2605.30792
HF Paper page: https://huggingface.co/papers/2605.30792… See the full description on the dataset page: https://huggingface.co/datasets/ayj111/openstbench-paired-set.moshi-paired-prefsynthetic_audio_paired_preferences
Synthetic Audio Paired Preferences
A synthetic spoken-dialogue DPO (Direct Preference Optimization) dataset with 8482 examples,
designed for preference alignment of textless speech language models.
Dataset Summary
Each example corresponds to a single assistant turn in a two-speaker conversation.
For every turn the dataset provides:
The spoken prompt (the last user utterance, as audio + text)
A chosen response: the best model-generated spoken reply (scored by an LLM… See the full description on the dataset page: https://huggingface.co/datasets/vprak17/synthetic_audio_paired_preferences.
