CoolFace
Datasetpublicgated

mlinmg/translation_dataset

translation_dataset Synthetic, expressive, multilingual speech for cross-lingual dubbing research. Each example pairs a style-annotated text with generated audio that clones an English reference voice: the voice stays the same, the language changes. ~1.9M examples in the one_speaker config 17 languages ~900 distinct reference speakers Samples Each sample shows the generated audio followed by the English reference voice that conditioned it. English… See the full description on the dataset page: https://huggingface.co/datasets/mlinmg/translation_dataset.

sourceHugging Faceupdated 3h agoView on Hugging Face
0likes729downloads

mlinmg/translation_dataset · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.

mlinmg/translation_dataset · CoolFace