CoolFace
Datasetpublic

eQOURSE/multilingual-speech

Multilingual Indian Conversational Speech A dataset of naturalistic, spontaneous two-speaker conversations across 13 Indian languages, with segment-level transcripts, speaker profiles, timestamps, and recording metadata. Designed for ASR, TTS, speaker diarization, and conversational speech research. Languages (13) Assamese, Bengali, Gujarati, Hindi, Kannada, Malayalam, Marathi, Nepali, Odia, Punjabi, Tamil, Telugu, Urdu. Content Conversations… See the full description on the dataset page: https://huggingface.co/datasets/eQOURSE/multilingual-speech.

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
0likes87downloads

eQOURSE/multilingual-speech · main · files are served by the source, never re-hosted here