instinct1912/multilingual-tts-selected-speech
Multilingual TTS Selected Speech This dataset contains text-and-audio pairs selected for TTS training. The initial audio/uzbek split has 2,184 utterances. Other languages and additional rows can be added as separate, append-only shards. Each row retains the complete structured text specification alongside audio and the OmniVoice fields utt, text, spk, and instruct. tts_tagged_text preserves the local delivery and vocal-event controls separately from the clean text field. For… See the full description on the dataset page: https://huggingface.co/datasets/instinct1912/multilingual-tts-selected-speech.
This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.
