sajalmadan0909/hindi_and_english_stt_tts_master_data
Hindi and English STT/TTS Master Data Combined speech dataset for Hindi and Indian English automatic speech recognition (ASR) and text-to-speech (TTS) training. Parquet shards embed WAV audio bytes with transcripts. Dataset structure hindi/<source>/train-*.parquet english/<source>/train-*.parquet Each config loads one source independently (~3.24M total rows, ~1.9 TB). Features Column Type Description audio Audio WAV bytes embedded in… See the full description on the dataset page: https://huggingface.co/datasets/sajalmadan0909/hindi_and_english_stt_tts_master_data.
018
No commit history came back for main. The revision may not exist, or the source declined the request.
