CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01manavtabbly /hindi_audio_dataset_testaudion<1K0 likes4.4k downloads11mo agoHugging Face02SPRINGLab /IndicTTS-Hindi Hindi Indic TTS Dataset This dataset is derived from the Indic TTS Database project, specifically using the Hindi monolingual recordings from both male and female speakers. The dataset contains high-quality speech recordings with corresponding text transcriptions, making it suitable for text-to-speech (TTS) research and development. Dataset Details Language: Hindi Total Duration: ~10.33 hours (Male: 5.16 hours, Female: 5.18 hours) Audio Format: WAV Sampling Rate: 48000Hz… See the full description on the dataset page: https://huggingface.co/datasets/SPRINGLab/IndicTTS-Hindi.audiotext-to-speech10K<n<100K36 likes1.3k downloads2y agoHugging Face03Ritwika03 /syspin_hindi_mergedaudio10K<n<100K1 likes686 downloads1y agoHugging Face04SPRINGLab /Hindi-1482Hrsaudio100K<n<1M5 likes522 downloads2y agoHugging Face05SPRINGLab /IndicVoices-R_Hindiaudiotext-to-speech10K<n<100K11 likes489 downloads2y agoHugging Face06Ritwika03 /hindi_karya_mergedaudio100K<n<1M0 likes399 downloads1y agoHugging Face07Paytmlabs /S2R_Shrutilipi_hindi Paytmlabs/S2R_Shrutilipi_hindi Hindi speech dataset prepared from ai4bharat/Shrutilipi for Ultravox training. Viewing samples on Hugging Face The hindi config stores audio inside Parquet. The website dataset viewer often cannot decode that and shows no rows. To inspect examples in the browser, open the Subset (config) drop-down and choose hindi_text_samples — text and continuation only (~2000 rows). Ultravox training should keep using subset hindi (full audio).… See the full description on the dataset page: https://huggingface.co/datasets/Paytmlabs/S2R_Shrutilipi_hindi.audioautomatic-speech-recognition100K<n<1M0 likes370 downloads6mo agoHugging Face08collabora /hindi-asr-wdsaudio1M<n<10M0 likes352 downloads1y agoHugging Face09MatrixSpeechAI /All_Hindi_ASR_v1.1audio10K<n<100K0 likes310 downloads2y agoHugging Face10MatrixSpeechAI /All_Hindi_ASR_v1.2audio10K<n<100K0 likes293 downloads2y agoHugging Face11nameissakthi /hindi-english-bilingual Hindi/English/Hinglish Bilingual TTS Dataset Synthetic TTS dataset generated by Rani voice (ai4bharat/indic-parler-tts) for training a lightweight bilingual student TTS model. Designed for natural-sounding Hindi, English, and Hinglish (code-switched) speech synthesis. Dataset Summary Property Value Total utterances 23,277 Total audio ~4.7GB (24kHz WAV) Languages Hindi (hi), English (en), Hinglish (bi) Sample rate 24kHz Voice Rani —… See the full description on the dataset page: https://huggingface.co/datasets/nameissakthi/hindi-english-bilingual.audiotext-to-speech10K<n<100K0 likes288 downloads6mo agoHugging Face12Ritwika03 /hindi_indic_voice_raudio10K<n<100K0 likes287 downloads1y agoHugging Face13Sheeba2026 /bharatvani-hindi-speech-corpusgated BharatVani Hindi Speech Corpus (150-Hour Studio Dataset) Proprietary Speech Asset • TheCreatorOS • BharatVani AI 1. Overview The BharatVani Hindi Speech Corpus is an enterprise-grade, high-fidelity Indian speech dataset engineered specifically for training sovereign neural Text-to-Speech (TTS) models, voice cloning engines, and speech foundation models in Devanagari Hindi. Audio Clips: 103,784 Verified Studio Audio Clips (24,000 Hz, 16-bit Mono… See the full description on the dataset page: https://huggingface.co/datasets/Sheeba2026/bharatvani-hindi-speech-corpus.audiotext-to-speech100K<n<1M1 likes284 downloads5d agoHugging Face14SayantanJoker /All_Hindi_ASR_v1.1audio10K<n<100K0 likes251 downloads1y agoHugging Face15SayantanJoker /processed_seamless_align_hindi_chunk_3audio10K<n<100K0 likes210 downloads1y agoHugging Face16SayantanJoker /processed_seamless_align_hindi_chunk_1audio10K<n<100K0 likes205 downloads1y agoHugging Face17backpropSukuna /hindi-audio-stories-20-30s Hindi Audio Stories — 20–30 s clips (Qwen3-ASR, denoised) Paired (audio, text) Hindi speech dataset. Each ~20–30 s denoised clip has its transcript in two scripts, stored as separate rows (script column): devanagari (Hindi) and latin (Hinglish romanization, uroman). ⚠️ Adult (NSFW) content. Research / non-commercial. Stats ~12.6k clips × 2 scripts ≈ 25k rows · ~94 h · mean 26.8 s (97% in 20–30 s) Audio: 24 kHz mono FLAC, UVR vocal-isolated (Mel-Band RoFormer —… See the full description on the dataset page: https://huggingface.co/datasets/backpropSukuna/hindi-audio-stories-20-30s.audioautomatic-speech-recognition10K<n<100K0 likes199 downloads4mo agoHugging Face18SayantanJoker /All_Hindi_ASR_Male_v1.1audio10K<n<100K0 likes197 downloads1y agoHugging Face19AniBirage /HindiDownloadedData2audio10K<n<100K0 likes195 downloads2y agoHugging Face20SayantanJoker /processed_seamless_align_hindi_chunk_5audio10K<n<100K0 likes194 downloads1y agoHugging Face21En1gma02 /hindi_speech_10haudio10K<n<100K0 likes182 downloads2y agoHugging Face22SayantanJoker /processed_seamless_align_hindi_chunk_2audio10K<n<100K0 likes176 downloads1y agoHugging Face23Paytmlabs /S2R_Kathbhat_hindi Paytmlabs/S2R_Kathbhat_hindi Hindi speech dataset prepared from ai4bharat/Kathbath for Ultravox training. Schema Column Type Description audio Audio Speech audio text string Verbatim transcript continuation string LLM-generated continuation (≤50 words) Progress Train chunks: 19/19 Validation: done audio10K<n<100K0 likes175 downloads6mo agoHugging Face24SayantanJoker /processed_seamless_align_hindi_chunk_4audio10K<n<100K0 likes174 downloads1y agoHugging Face25equal-ai /merged-hindi-audio-datasetaudio10K<n<100K0 likes171 downloads7mo agoHugging Face26Nipurn /orpheus-tts-hindiaudio10K<n<100K0 likes166 downloads4mo agoHugging Face27RutwikShete /hindi_dataset_stats_catagorical_description_audioaudio100K<n<1M1 likes154 downloads2y agoHugging Face28SayantanJoker /processed_seamless_align_hindi_chunk_6audio10K<n<100K0 likes154 downloads1y agoHugging Face29SayantanJoker /Shrutilipi_Hindi_resampled_44100_merged_10audio10K<n<100K0 likes146 downloads1y agoHugging Face30ArchCoder /hindi-whisper-chunks Hindi Whisper Chunks Preprocessed, feature-extracted audio chunks and labels used to fine-tune ArchCoder/whisper-small-hindi-lora, a LoRA adaptation of Whisper-small for Hindi speech recognition. Dataset Summary Raw Hindi audio recordings (approximately 12 minutes each) were segmented into short, Whisper-compatible chunks and converted into model-ready features. This dataset is the output of that preprocessing pipeline: Whisper-format log-mel filterbank features… See the full description on the dataset page: https://huggingface.co/datasets/ArchCoder/hindi-whisper-chunks.audio1K<n<10K0 likes142 downloads19d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.