CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01deepdml /microsoft-speech-corpus-indian Microsoft Speech Corpus – Indian Languages Dataset Description This dataset is a redistribution of the Microsoft Speech Corpus (Indian Languages) containing conversational and phrasal speech training and test data for Telugu, Tamil, and Gujarati languages. Each entry includes an audio recording and its corresponding transcript. Attribution required: "Data provided by Microsoft and SpeechOcean.com" ⚠️ License: This data is provided for research purposes only. Commercial… See the full description on the dataset page: https://huggingface.co/datasets/deepdml/microsoft-speech-corpus-indian.audioautomatic-speech-recognition100K<n<1M3 likes306 downloads7mo agoHugging Face02satwc-reddy /indian-language-deepfake-speech Multilingual Deepfake Speech Dataset (Indian Languages) Overview This dataset contains real and synthetic speech across four Indian languages: Telugu Tamil Malayalam Konkani It is designed for deepfake speech detection and cross-language generalization research. Composition Real speech: OpenSLR datasets Synthetic speech: MMS-TTS (neural TTS) SIGVC (signal-based transformations) RVC (voice conversion) Total samples: ~15,000+ Average duration: ~5–7… See the full description on the dataset page: https://huggingface.co/datasets/satwc-reddy/indian-language-deepfake-speech.audio-classification0 likes97 downloads5mo agoHugging Face03humyn-labs /Indian-Emotional-Speech-Corpus Indian Emotional Speech Corpus Dataset Description This dataset comprises high-quality audio recordings of Indian speakers reading a standardized 50-word paragraph in four distinct emotional tones — happy, sad, surprised, and angry. Each recording is approximately 20–25 seconds long and includes the full paragraph with tone shifts at specific points. Text spoken by all participants: (happy tone) Last Monday was perfect—I got the job I’d been dreaming of! I screamed… See the full description on the dataset page: https://huggingface.co/datasets/humyn-labs/Indian-Emotional-Speech-Corpus.audioaudio-classificationn<1K7 likes67 downloads7mo agoHugging Face04sudhanshu12 /KAI-indian-emotional-speech-corpus Indian Emotional Speech Corpus Dataset Description This dataset comprises high-quality audio recordings of Indian speakers reading a standardized 50-word paragraph in four distinct emotional tones — happy, sad, surprised, and angry. Each recording is approximately 20–25 seconds long and includes the full paragraph with tone shifts at specific points. Text spoken by all participants: (happy tone) Last Monday was perfect—I got the job I’d been dreaming of! I screamed… See the full description on the dataset page: https://huggingface.co/datasets/sudhanshu12/KAI-indian-emotional-speech-corpus.audioaudio-classificationn<1K0 likes50 downloads8mo agoHugging Face05learn-abc /indian-speech-audio-extendedaudion<1K0 likes13 downloads1y agoHugging Face06theothertom /indian_speech_audiotextn<1K0 likes9 downloads3y agoHugging Face07DataoceanAI /Indian_English_Speech_Recognition_Corpus_Conversations ID King-ASR-631 Language English Duration 200 hours Speakers 200 People Parameters 16kHz, 16bits Recording Device Mobile URL https://dataoceanai.com/datasets/asr/indian-english-speech-recognition-corpus-conversations-mobile/ 0 likes8 downloads2y agoHugging Face08manishchandraguturu /Indian_english_Example_speech0 likes1 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.