datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
data-voice-vietnamese-restaurant-quan-oc
Vietnamese Restaurant Order Speech
This dataset contains Vietnamese spoken restaurant orders paired with text transcripts. Each utterance typically includes a table number, item quantities, dishes, drinks, and add-ons.
Dataset Structure
Files are split into subdirectories by filename-derived speaker_code to satisfy Hugging Face repository file-count limits:
metadata.csv: one row per audio sample.
audio/{speaker_code}/*.wav: mono WAV audio files.… See the full description on the dataset page: https://huggingface.co/datasets/EmilyNguyen235/data-voice-vietnamese-restaurant-quan-oc.yt4_chunked_speech_restorised_tts_train
yt4_chunked_speech_restorised_tts_train
This is a gated Russian TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: ru (Russian)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt4_chunked_speech_restorised_tts_train.yt_chunked_speech_restorised_tts_train
yt_chunked_speech_restorised_tts_train
This is a gated Russian TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: ru (Russian)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt_chunked_speech_restorised_tts_train.yt3_chunked_speech_restorised_tts_train
yt3_chunked_speech_restorised_tts_train
This is a gated Russian TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: ru (Russian)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes
Prepared for TTS… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt3_chunked_speech_restorised_tts_train.yt1_chunked_speech_restorised_tts_train
yt1_chunked_speech_restorised_tts_train
This is a gated Russian TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: ru (Russian)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes
Prepared for TTS… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt1_chunked_speech_restorised_tts_train.yt2_chunked_speech_restorised_tts_train
yt2_chunked_speech_restorised_tts_train
This is a gated Russian TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: ru (Russian)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes
Prepared for TTS… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt2_chunked_speech_restorised_tts_train.miscellaneous_yt_chunked_speech_restorised_tts_train
miscellaneous_yt_chunked_speech_restorised_tts_train
This is a gated Uzbek TTS training dataset from instinct-org.
This repository contains tokenized or prepared speech data for text-to-speech training workflows.
Language
Primary language: uz (Uzbek)
Intended Use
text-to-speech training
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license review
Data Notes
Prepared for… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/miscellaneous_yt_chunked_speech_restorised_tts_train.dataset-from-restorecv-corpus21_be-sidon-restored-1000
cv-corpus21_be-sidon-restored-1000
Прыклад датасэта з 639 запісамі (Belarusian, Common Voice validated),
дзе audio — адноўлены WAV (48 кГц), original_audio — арыгінальны кліп,
а таксама sentence і speaker.
Створана: 2025-09-26.
ataturk_voice_no_restorationrestaurant_order_HSR_test
Dataset Card for "restaurant_order_HSR_test"
More Information needed
tbp_chunked_speech_restorised
tbp_chunked_speech_restorised
This is a gated Russian speech-restorised chunked speech dataset from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: ru (Russian)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and license… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/tbp_chunked_speech_restorised.darija-restaurant-audiobe-sidon-restored-sample-10-fixed
be-sidon-restored-sample-10-fixed
Прыклад датасэта з 10 запісамі (Belarusian, Common Voice validated),
дзе audio — адноўлены WAV (48 кГц), original_audio — арыгінальны кліп,
а таксама sentence і speaker.
Створана: 2025-09-25.
whisper_restaurant_trainingespeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs
espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/espeech_podcasts_chunked_speech_restorised_tts_train_clone_pairs.default_voices_chunked_speech_restorised
default_voices_chunked_speech_restorised
This is a gated Uzbek speech-restorised chunked speech dataset from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: uz (Uzbek)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_speech_restorised.cv_chunked_speech_restorised
cv_chunked_speech_restorised
This is a gated Uzbek speech-restorised chunked speech dataset from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: uz (Uzbek)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/cv_chunked_speech_restorised.default_voices_chunked_speech_restorised_tts_train_clone_pairs
default_voices_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Uzbek TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: uz (Uzbek)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/default_voices_chunked_speech_restorised_tts_train_clone_pairs.yt2_chunked_speech_restorised_tts_train_clone_pairs
yt2_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt2_chunked_speech_restorised_tts_train_clone_pairs.yt3_chunked_speech_restorised_tts_train_clone_pairs
yt3_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt3_chunked_speech_restorised_tts_train_clone_pairs.restaurant_order_local_test_colab
Dataset Card for "restaurant_order_local_test_colab"
More Information needed
ataturk_voice_restoratedaudiobook_chunked_speech_restorised
audiobook_chunked_speech_restorised
This is a gated Uzbek speech-restorised chunked speech dataset from instinct-org.
This repository contains speech audio and transcripts for speech-to-text training, evaluation, or data preparation workflows.
Language
Primary language: uz (Uzbek)
Intended Use
speech-to-text training and evaluation
Internal dataset curation, quality checks, and model evaluation
Research or commercial use only after access approval and… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/audiobook_chunked_speech_restorised.tbp_chunked_speech_restorised_tts_train_clone_pairs
tbp_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/tbp_chunked_speech_restorised_tts_train_clone_pairs.audiobook_chunked_speech_restorised_tts_train_clone_pairs
audiobook_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Uzbek TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: uz (Uzbek)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/audiobook_chunked_speech_restorised_tts_train_clone_pairs.miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs
miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Uzbek TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: uz (Uzbek)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json:… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/miscellaneous_yt_chunked_speech_restorised_tts_train_clone_pairs.yt4_chunked_speech_restorised_tts_train_clone_pairs
yt4_chunked_speech_restorised_tts_train_clone_pairs
This is a gated Russian TTS training clone-pair dataset.
It contains tokenized speaker-reference and target pairs for text-to-speech voice adaptation workflows.
Language
Primary language: ru (Russian)
Contents
audios/shard-*.tar: tokenized audio shards
txts/shard-*.jsonl: per-example metadata and text fields
data.lst: repository-relative shard manifest
tokenized_dataset_summary.json: upload-time… See the full description on the dataset page: https://huggingface.co/datasets/instinct-org/yt4_chunked_speech_restorised_tts_train_clone_pairs.cv-corpus20_be-sidon-restored-1000
cv-corpus20_be-sidon-restored-1000
Прыклад датасэта з 1000 запісамі (Belarusian, Common Voice validated),
дзе audio — адноўлены WAV (48 кГц), original_audio — арыгінальны кліп,
а таксама sentence і speaker.
Створана: 2025-09-26.
whisper_indian_restaurant_training
