CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01humair025 /Urdu-ONYX-WAV-kanade-Annotated Urdu-ONYX-WAV-real-Annotated Enhanced version of Urdu-ONYX-WAV-real with phoneme annotations and Kanade tokenizer features. Dataset Statistics Total Samples: 26,217 Total Duration: 42.77 hours Average Duration: 5.87 seconds Duration Range: 0.65s - 122.23s Average Phonemes: 18.5 per sample Average Kanade Tokens: 151.1 per sample Global Embedding Dimension: 128 New Columns This dataset adds the following columns: duration (float): Audio duration in seconds… See the full description on the dataset page: https://huggingface.co/datasets/humair025/Urdu-ONYX-WAV-kanade-Annotated.tabulartext-to-speech100K<n<1M0 likes1.3k downloads8mo agoHugging Face02imvladikon /hebrew_speech_kan Dataset Card for Dataset Name Dataset Summary Hebrew Dataset for ASR Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances {'audio': {'path': '/root/.cache/huggingface/datasets/downloads/extracted/8ce7402f6482c6053251d7f3000eec88668c994beb48b7ca7352e77ef810a0b6/train/e429593fede945c185897e378a5839f4198.wav', 'array': array([-0.00265503, -0.0018158… See the full description on the dataset page: https://huggingface.co/datasets/imvladikon/hebrew_speech_kan.audioautomatic-speech-recognition10K<n<100K14 likes518 downloads3y agoHugging Face03lilgoose777 /kantipur-interview-data3gated Nepali Speech Dataset (YouTube-sourced) 83 labeled speech segments, split by channel (not by individual video) so the same speaker/recording can't appear in more than one split. Splits train: 83 segments validation: 0 segments test: 0 segments Transcript columns — read this before training Each segment carries three transcript variants. They are NOT interchangeable: text_original — the YouTube caption text (if any) that overlapped this segment's… See the full description on the dataset page: https://huggingface.co/datasets/lilgoose777/kantipur-interview-data3.audioautomatic-speech-recognitionn<1K0 likes29 downloads28d agoHugging Face04lilgoose777 /kantipur-interview-data1gated Nepali Speech Dataset (YouTube-sourced) 204 labeled speech segments, split by channel (not by individual video) so the same speaker/recording can't appear in more than one split. Splits train: 204 segments validation: 0 segments test: 0 segments Transcript columns — read this before training Each segment carries three transcript variants. They are NOT interchangeable: text_original — the YouTube caption text (if any) that overlapped this… See the full description on the dataset page: https://huggingface.co/datasets/lilgoose777/kantipur-interview-data1.audioautomatic-speech-recognitionn<1K0 likes28 downloads28d agoHugging Face05lilgoose777 /kantipur-interview-data2gated Nepali Speech Dataset (YouTube-sourced) 157 labeled speech segments, split by channel (not by individual video) so the same speaker/recording can't appear in more than one split. Splits train: 157 segments validation: 0 segments test: 0 segments Transcript columns — read this before training Each segment carries three transcript variants. They are NOT interchangeable: text_original — the YouTube caption text (if any) that overlapped this… See the full description on the dataset page: https://huggingface.co/datasets/lilgoose777/kantipur-interview-data2.audioautomatic-speech-recognitionn<1K0 likes26 downloads28d agoHugging Face06ShimogaAIteam /conversational_kannada_stt Conversational Kannada STT This dataset contains corrected transcriptions of conversational Kannada speech,prepared specifically for fine-tuning Whisper models on conversational and dialectal Kannada. Unlike many ASR datasets, this release provides pre-computed Whisper input features (log-Mel spectrograms)so you can train/fine-tune Whisper models without raw audio processing. Dataset Creation Source Audio: Publicly available YouTube videos in Kannada. Initial… See the full description on the dataset page: https://huggingface.co/datasets/ShimogaAIteam/conversational_kannada_stt.textautomatic-speech-recognition1K<n<10K1 likes22 downloads1y agoHugging Face07lilgoose777 /kantipur-news-datagated Nepali Speech Dataset (YouTube-sourced) 74 labeled speech segments, split by channel (not by individual video) so the same speaker/recording can't appear in more than one split. Splits train: 74 segments validation: 0 segments test: 0 segments Transcript columns — read this before training Each segment carries three transcript variants. They are NOT interchangeable: text_original — the YouTube caption text (if any) that overlapped this segment's… See the full description on the dataset page: https://huggingface.co/datasets/lilgoose777/kantipur-news-data.audioautomatic-speech-recognitionn<1K0 likes22 downloads28d agoHugging Face08Speech-data /Kannada-Speech-Dataset 🎧 Kannada Speech Dataset The Kannada Speech Dataset is a high-quality speech audio dataset designed to deliver structured and reliable audio data for AI and machine learning workflows. It includes 90 hours of audio data across 651 files, available in MP3 and WAV formats, with a total size of 220 MB. This well-organized audio dataset provides balanced and representative voice data, with 48% female and 52% male speakers, and an age range spanning from 18 to 50+ years. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Kannada-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes17 downloads6mo agoHugging Face09Kannada-LLM-Labs /Fleurs-KnThis is a filtered version of the Fleurs dataset only containing samples of Kannada language. The dataset contains total of 2283 training, 368 validation and 838 test samples. Data Sample: {'id': 1053, 'num_samples': 226560, 'path': '/home/ravi.naik/.cache/huggingface/datasets/downloads/extracted/e7c8b501d4e6892673b6dc291d42de48e7987b0d2aa6471066a671f686224ed1/10000267636955490843.wav', 'audio': {'path': 'train/10000267636955490843.wav', 'array': array([ 0. , 0.… See the full description on the dataset page: https://huggingface.co/datasets/Kannada-LLM-Labs/Fleurs-Kn.audioautomatic-speech-recognition1K<n<10K0 likes15 downloads3y agoHugging Face10humair025 /Urdu-ONYX-WAV-kanade-V2gated Urdu-ONYX-WAV-kanade-Annotated-V2 Version 2.0 - Artifact-Free Edition 🎉 Overview This is an improved version of the Urdu-ONYX-WAV dataset, tokenized with the Kanade neural codec and optimized for artifact-free audio decoding. This dataset contains 143,627 samples of high-quality Urdu speech with comprehensive linguistic and acoustic annotations, totaling ~244 hours (~10 days) of continuous audio. Key Features 🎯 Large-Scale: 143K+ samples, 244+ hours of… See the full description on the dataset page: https://huggingface.co/datasets/humair025/Urdu-ONYX-WAV-kanade-V2.tabulartext-to-speech100K<n<1M0 likes3 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.