clone_pairs
default_voices_chunked_speech_restorised_tts_train_clone_pairs_raw
default_voices_chunked_speech_restorised_tts_train_clone_pairs_raw
This is a gated Uzbek raw TTS clone-pair plan for default_voices_chunked_speech_restorised.
It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards.
Contents
clone_pair_plan.jsonl: one reference/target pair per row
clone_pair_plan_summary.json: upload-time summary
Dataset Summary
Field
Value
Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_speech_restorised_tts_train_clone_pairs_raw.yt3_chunked_speech_restorised_tts_train_clone_pairs
yt3_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt3_chunked_speech_restorised_tts_train_clone_pairs.yt2_chunked_speech_restorised_tts_train_clone_pairs
yt2_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt2_chunked_speech_restorised_tts_train_clone_pairs.tbp_chunked_speech_restorised_tts_train_clone_pairs
tbp_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/tbp_chunked_speech_restorised_tts_train_clone_pairs.espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_raw
espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_raw
This is a gated Russian raw TTS clone-pair plan for espeech_podcasts_chunked_speech_restorised.
It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards.
Contents
clone_pair_plan.jsonl: one reference/target pair per row
clone_pair_plan_summary.json: upload-time summary
Dataset Summary
Field
Value
Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_raw.audiobook_chunked_speech_restorised_tts_train_clone_pairs
audiobook_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Uzbek TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: uz (Uzbek)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/audiobook_chunked_speech_restorised_tts_train_clone_pairs.
