CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01parambharat /kannada_asr_corpusThe corpus contains roughly 360 hours of audio and transcripts in Kannada language. The transcripts have beed de-duplicated using exact match deduplication.automatic-speech-recognition100K<n<1M0 likes29 downloads4y agoHugging Face02InfoBayAI /Kannada_Call_Center_Audio_Dataset_Dual_ChannelgatedDataset Description: This dataset is a large-scale collection of 16,549 hours of processed Kannada (KN) dual-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems. It consists of real-world customer and agent speech recordings collected from call center environments. The dataset is organized in a dual-channel format… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Kannada_Call_Center_Audio_Dataset_Dual_Channel.audioautomatic-speech-recognitionn<1K0 likes24 downloads9d agoHugging Face03ShimogaAIteam /conversational_kannada_stt Conversational Kannada STT This dataset contains corrected transcriptions of conversational Kannada speech,prepared specifically for fine-tuning Whisper models on conversational and dialectal Kannada. Unlike many ASR datasets, this release provides pre-computed Whisper input features (log-Mel spectrograms)so you can train/fine-tune Whisper models without raw audio processing. Dataset Creation Source Audio: Publicly available YouTube videos in Kannada. Initial… See the full description on the dataset page: https://huggingface.co/datasets/ShimogaAIteam/conversational_kannada_stt.textautomatic-speech-recognition1K<n<10K1 likes22 downloads1y agoHugging Face04InfoBayAI /Kannada-Call-Center-Audio-Dataset-Single-ChannelgatedDataset Description: This dataset is a large-scale collection of 16,549 hours of processed Kannada (KN) single-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems. The dataset captures authentic speech characteristics such as tone variation, pauses, silence patterns, and natural speaking behaviour commonly observed in… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Kannada-Call-Center-Audio-Dataset-Single-Channel.audioautomatic-speech-recognitionn<1K0 likes20 downloads9d agoHugging Face05Speech-data /Kannada-Speech-Dataset 🎧 Kannada Speech Dataset The Kannada Speech Dataset is a high-quality speech audio dataset designed to deliver structured and reliable audio data for AI and machine learning workflows. It includes 90 hours of audio data across 651 files, available in MP3 and WAV formats, with a total size of 220 MB. This well-organized audio dataset provides balanced and representative voice data, with 48% female and 52% male speakers, and an age range spanning from 18 to 50+ years. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Kannada-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes17 downloads6mo agoHugging Face06Kannada-LLM-Labs /Fleurs-KnThis is a filtered version of the Fleurs dataset only containing samples of Kannada language. The dataset contains total of 2283 training, 368 validation and 838 test samples. Data Sample: {'id': 1053, 'num_samples': 226560, 'path': '/home/ravi.naik/.cache/huggingface/datasets/downloads/extracted/e7c8b501d4e6892673b6dc291d42de48e7987b0d2aa6471066a671f686224ed1/10000267636955490843.wav', 'audio': {'path': 'train/10000267636955490843.wav', 'array': array([ 0. , 0.… See the full description on the dataset page: https://huggingface.co/datasets/Kannada-LLM-Labs/Fleurs-Kn.audioautomatic-speech-recognition1K<n<10K0 likes15 downloads3y agoHugging Face07InfoBayAI /Kannada_Podcast_Audio_Dataset_Dual_Channelgated Dataset Description This dataset is a large-scale collection of 3,970 hours of processed Kannada dual-channel podcast audio recordings, containing 57,569 hours of processed podcast audio recordings across 12 languages, designed to support the development and training of advanced speech AI and conversational AI systems. It captures real-world podcast conversations across diverse topics and formats. The dataset is organized in a dual-channel format, where corresponding speaker… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Kannada_Podcast_Audio_Dataset_Dual_Channel.audioautomatic-speech-recognitionn<1K0 likes15 downloads9d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.