CoolFace
Datasetpublicgated

sajalmadan0909/hindi_and_english_stt_tts_codemix_data

Hindi and English STT/TTS Codemix Data Hinglish (Hindi-English code-mixed) speech dataset for automatic speech recognition (ASR) and text-to-speech (TTS) research. Dataset Description Each row is a timestamped speech segment clipped from conversational Hinglish audio recordings. Column Type Description text string Transcript of the speech segment (Hinglish) audio audio (16 kHz mono) Corresponding audio clip duration float32 Clip duration in seconds… See the full description on the dataset page: https://huggingface.co/datasets/sajalmadan0909/hindi_and_english_stt_tts_codemix_data.

sourceHugging Faceotherupdated 3mo agoView on Hugging Face
0likes12downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.