datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
malayalam-tts-pro-voice
Malayalam Speech Dataset (Text + Audio)
This dataset contains Malayalam speech audio clips paired with text transcripts.It is designed for training and fine-tuning ASR (Automatic Speech Recognition),TTS (Text-to-Speech) models, and speech-to-speech translation systems.
Emotions Added
giggles
laughs
long pause
chuckles
whispers
gasps
clears throat
singing
laughs nervously
burps
exhales
📁 Dataset Structure
Column
Description
audio
Path to the… See the full description on the dataset page: https://huggingface.co/datasets/sachin6624/malayalam-tts-pro-voice.kokoro-dialogue-asr
Kokoro dialogue ASR set
500 single-speaker English clips of ~25-30 s, synthesized with
Kokoro-82M (voice af_heart) reading
procedurally generated spoken-monologue passages. Intended as augmentation for ASR
fine-tuning, not as a standalone training set.
split
clips
hours
train
406
3.14
test
94
0.72
Fields
audio — 16 kHz mono
transcription — orthographic transcript, exactly the text that was synthesized
topic — which passage template produced it… See the full description on the dataset page: https://huggingface.co/datasets/sachin6624/kokoro-dialogue-asr.audio-data
