CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01marcosremar2 /minimind-ptbr-kokoro-tts-68k MiniMind PT-BR Kokoro TTS 68k Synthetic Brazilian Portuguese speech dataset generated for the MiniMind/Tucano2 -> Mimi Talker fine-tuning experiments. Contents audio/: 68,486 WAV files generated with Kokoro TTS. manifests/kokoro_groq25k_audio_manifest.jsonl: source text manifest. manifests/kokoro_groq25k_audio_manifest_with_wavs.jsonl: manifest with WAV paths. manifests/kokoro_groq25k_audio_manifest_with_mimi.jsonl: manifest with Mimi token references.… See the full description on the dataset page: https://huggingface.co/datasets/marcosremar2/minimind-ptbr-kokoro-tts-68k.audiotext-to-speech1K<n<10K0 likes168 downloads4mo agoHugging Face02sachin6624 /kokoro-dialogue-asr Kokoro dialogue ASR set 500 single-speaker English clips of ~25-30 s, synthesized with Kokoro-82M (voice af_heart) reading procedurally generated spoken-monologue passages. Intended as augmentation for ASR fine-tuning, not as a standalone training set. split clips hours train 406 3.14 test 94 0.72 Fields audio — 16 kHz mono transcription — orthographic transcript, exactly the text that was synthesized topic — which passage template produced it… See the full description on the dataset page: https://huggingface.co/datasets/sachin6624/kokoro-dialogue-asr.audioautomatic-speech-recognitionn<1K0 likes17 downloads2mo agoHugging Face03speed-tb /kokborokgated Speed-Tb Phase 1 Kokborok Narration Dataset Description The Kok Borok Speech Dataset, developed as part of the Speech Datasets and Models for Tibeto-Burman Languages (Project SpeeD-TB), funded under Mission Bhashini, is a transcribed speech corpus of the language. The full dataset comprises over 200 hours of high-quality audio recordings paired with accurate transcriptions in both IPA and Roman script, making it ** one of the largest speech resources for the… See the full description on the dataset page: https://huggingface.co/datasets/speed-tb/kokborok.audioautomatic-speech-recognitionn<1K0 likes17 downloads15d agoHugging Face04french-datasets /sdialog_voices-kokoroCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données sdialog/voices-kokoro. automatic-speech-recognition0 likes10 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.