CoolFace
21 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01instinct-org /default_voices_chunked_speech_restorised_tts_train_clone_pairs_rawgated default_voices_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Uzbek raw TTS clone-pair plan for default_voices_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes10 downloads4mo agoHugging Face02instinct-org /yt3_chunked_speech_restorised_tts_train_clone_pairsgated yt3_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt3_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes10 downloads4mo agoHugging Face03instinct-org /yt2_chunked_speech_restorised_tts_train_clone_pairsgated yt2_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt2_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes9 downloads4mo agoHugging Face04instinct-org /tbp_chunked_speech_restorised_tts_train_clone_pairsgated tbp_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/tbp_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes8 downloads4mo agoHugging Face05instinct-org /espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_rawgated espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for espeech_podcasts_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes8 downloads4mo agoHugging Face06instinct-org /audiobook_chunked_speech_restorised_tts_train_clone_pairsgated audiobook_chunked_speech_restorised_tts_train_clone_pairs This is a gated Uzbek TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: uz (Uzbek) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/audiobook_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes8 downloads4mo agoHugging Face07instinct-org /espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairsgated espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes8 downloads4mo agoHugging Face08instinct-org /yt4_chunked_speech_restorised_tts_train_clone_pairsgated yt4_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt4_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes8 downloads4mo agoHugging Face09instinct-org /all15_speaker_deduped_tts_train_clone_pairs_rawgated All-15 Speaker-Deduped TTS Train Clone Pairs Raw This dataset contains raw metadata rows for speaker-deduped TTS voice-clone training pairs. It does not contain audio bytes. Rows point back to source audio records and include reference/target metadata, language, dataset, tier, and precomputed speaker-similarity fields from the mining pipeline. Contents data/train/distinct_speaker_clone_pair_plan.jsonl.gz: all survivor rows. data/by_dataset/*.jsonl.gz: the same… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/all15_speaker_deduped_tts_train_clone_pairs_raw.tabulartext-to-speech10K<n<100K0 likes8 downloads4mo agoHugging Face10instinct-org /tbp_chunked_speech_restorised_tts_train_clone_pairs_rawgated tbp_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for tbp_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key tbp_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/tbp_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes7 downloads4mo agoHugging Face11instinct-org /default_voices_chunked_speech_restorised_tts_train_clone_pairsgated default_voices_chunked_speech_restorised_tts_train_clone_pairs This is a gated Uzbek TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: uz (Uzbek) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes7 downloads4mo agoHugging Face12instinct-org /miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs_rawgated miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Uzbek raw TTS clone-pair plan for miscellaneous_yt_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes7 downloads4mo agoHugging Face13instinct-org /yt3_chunked_speech_restorised_tts_train_clone_pairs_rawgated yt3_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for yt3_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key yt3_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt3_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes7 downloads4mo agoHugging Face14instinct-org /yt_chunked_speech_restorised_tts_train_clone_pairsgated yt_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes7 downloads4mo agoHugging Face15instinct-org /yt1_chunked_speech_restorised_tts_train_clone_pairs_rawgated yt1_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for yt1_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key yt1_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt1_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes6 downloads4mo agoHugging Face16instinct-org /yt2_chunked_speech_restorised_tts_train_clone_pairs_rawgated yt2_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for yt2_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key yt2_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt2_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes6 downloads4mo agoHugging Face17instinct-org /yt4_chunked_speech_restorised_tts_train_clone_pairs_rawgated yt4_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for yt4_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key yt4_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt4_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes6 downloads4mo agoHugging Face18instinct-org /miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairsgated miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs This is a gated Uzbek TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: uz (Uzbek) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes6 downloads4mo agoHugging Face19instinct-org /yt1_chunked_speech_restorised_tts_train_clone_pairsgated yt1_chunked_speech_restorised_tts_train_clone_pairs This is a gated Russian TTS training clone-pair dataset. It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows. Language Primary language: ru (Russian) Contents audios/shard-*.tar: tokenized audio shards txts/shard-*.jsonl: per-example metadata and text fields data.lst: repository-relative shard manifest tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt1_chunked_speech_restorised_tts_train_clone_pairs.audiotext-to-speech0 likes6 downloads4mo agoHugging Face20instinct-org /audiobook_chunked_speech_restorised_tts_train_clone_pairs_rawgated audiobook_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Uzbek raw TTS clone-pair plan for audiobook_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/audiobook_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes5 downloads4mo agoHugging Face21instinct-org /yt_chunked_speech_restorised_tts_train_clone_pairs_rawgated yt_chunked_speech_restorised_tts_train_clone_pairs_raw This is a gated Russian raw TTS clone-pair plan for yt_chunked_speech_restorised. It contains speaker-reference and target metadata rows used to build tokenized clone-pair training shards. Contents clone_pair_plan.jsonl: one reference/target pair per row clone_pair_plan_summary.json: upload-time summary Dataset Summary Field Value Source dataset key yt_chunked_speech_restorised… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt_chunked_speech_restorised_tts_train_clone_pairs_raw.text-to-speech0 likes4 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.