mlinmg/translation_dataset
translation_dataset Synthetic, expressive, multilingual speech for cross-lingual dubbing research. Each example pairs a style-annotated text with generated audio that clones an English reference voice: the voice stays the same, the language changes. ~1.9M examples in the one_speaker config 17 languages ~900 distinct reference speakers Samples Each sample shows the generated audio followed by the English reference voice that conditioned it. English… See the full description on the dataset page: https://huggingface.co/datasets/mlinmg/translation_dataset.
This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.
