datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
minimind-ptbr-kokoro-tts-68k
MiniMind PT-BR Kokoro TTS 68k
Synthetic Brazilian Portuguese speech dataset generated for the MiniMind/Tucano2 -> Mimi Talker fine-tuning experiments.
Contents
audio/: 68,486 WAV files generated with Kokoro TTS.
manifests/kokoro_groq25k_audio_manifest.jsonl: source text manifest.
manifests/kokoro_groq25k_audio_manifest_with_wavs.jsonl: manifest with WAV paths.
manifests/kokoro_groq25k_audio_manifest_with_mimi.jsonl: manifest with Mimi token references.… See the full description on the dataset page: https://huggingface.co/datasets/marcosremar2/minimind-ptbr-kokoro-tts-68k.kokoro-dialogue-asr
Kokoro dialogue ASR set
500 single-speaker English clips of ~25-30 s, synthesized with
Kokoro-82M (voice af_heart) reading
procedurally generated spoken-monologue passages. Intended as augmentation for ASR
fine-tuning, not as a standalone training set.
split
clips
hours
train
406
3.14
test
94
0.72
Fields
audio — 16 kHz mono
transcription — orthographic transcript, exactly the text that was synthesized
topic — which passage template produced it… See the full description on the dataset page: https://huggingface.co/datasets/sachin6624/kokoro-dialogue-asr.kokborok
Speed-Tb Phase 1 Kokborok Narration
Dataset Description
The Kok Borok Speech Dataset, developed as part of the
Speech Datasets and Models for Tibeto-Burman Languages (Project SpeeD-TB),
funded under Mission Bhashini, is a transcribed speech corpus of the language.
The full dataset comprises over 200 hours of high-quality audio recordings paired with accurate transcriptions in both IPA and Roman script, making it
** one of the largest speech resources for the… See the full description on the dataset page: https://huggingface.co/datasets/speed-tb/kokborok.sdialog_voices-kokoroCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données sdialog/voices-kokoro.
