CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ghanaopenai /navigation-corpus-speech-full-dagbani This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Ghana TTS Navigation Corpus — Dagbani Synthetic speech dataset for navigation. Structure audio/ – all .wav audio files text/ – matching .txt files with transcriptions metadata.csv – full metadata table audiotext-to-speech1K<n<10K0 likes313 downloads3mo agoHugging Face02ghanaopenai /ghana-female-twi-speech-asr-full-length This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Audio-text dataset with 76 pairs of Twi (Ghanaian language) speech data. Structure audio/ - WAV audio files ({len(pairs)} files) text/ - Corresponding text transcripts ({len(pairs)} files) dataset_manifest.json - Links audio to… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-full-length.audioautomatic-speech-recognitionn<1K0 likes220 downloads3mo agoHugging Face03ghanaopenai /navigation-corpus-speech-full-twi This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Ghana TTS Navigation Corpus — Twi Synthetic speech dataset for navigation. Structure audio/ – all .wav audio files text/ – matching .txt files with transcriptions metadata.csv – full metadata table audiotext-to-speech1K<n<10K0 likes197 downloads3mo agoHugging Face04Wi-Fi /korean-full-duplex-synthetic-dataset-preview Korean Full-Duplex Synthetic Dataset Preview Overview Public preview of a Korean full-duplex synthetic speech dataset. This repository contains 100 conversations sampled from a corpus of 89,273 conversations (2,000.5 hours); it does not publish the full corpus audio. Preview contents 100 conversation WAV files data/representative.jsonl 24 kHz, mono, 16-bit PCM Events: normal, barge_in, backchannel, cutoff_by_user Annotation format… See the full description on the dataset page: https://huggingface.co/datasets/Wi-Fi/korean-full-duplex-synthetic-dataset-preview.audioautomatic-speech-recognitionn<1K1 likes150 downloads1mo agoHugging Face05Codyfederer /tr-full-dataset TR-Full_dataset This is a merged speech dataset containing 41427 audio segments from 88 source datasets. Dataset Information Total Segments: 41427 Speakers: 222 Languages: tr Emotions: neutral, angry, sad, happy Original Datasets: 88 Dataset Structure Each example contains: audio: Audio file (WAV format, original sampling rate preserved) text: Transcription of the audio speaker_id: Unique speaker identifier (made unique across all merged… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/tr-full-dataset.audioautomatic-speech-recognition10K<n<100K6 likes143 downloads1y agoHugging Face06ghanaopenai /navigation-corpus-speech-full-ewe This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Ghana TTS Navigation Corpus — Ewe Synthetic speech dataset for navigation. Structure audio/ – all .wav audio files text/ – matching .txt files with transcriptions metadata.csv – full metadata table audiotext-to-speech1K<n<10K0 likes130 downloads3mo agoHugging Face07Trelis /VoxPopuli-Platinum-en-full VoxPopuli-Platinum-en-full Complete 142,066-row / ~404-hour English VoxPopuli Platinum dataset. Reach out to data@trelis.com to purchase access or discuss a larger custom-curation engagement. Training Results These results show why the Platinum labels matter. We compare the base model, fine-tuning on raw VoxPopuli transcripts, and fine-tuning on this internally filtered Platinum dataset. Evaluation uses the same english-spoken corpus WER setup across four… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/VoxPopuli-Platinum-en-full.audioautomatic-speech-recognition100K<n<1M0 likes115 downloads4mo agoHugging Face08MohammadGholizadeh /fleurs-farsi-fullaudioautomatic-speech-recognition1K<n<10K3 likes56 downloads1y agoHugging Face09Kukedlc /openslr61-es-ar-full openslr61-es-ar-full OpenSLR 61 (Crowdsourced high-quality Argentinian Spanish) consolidado COMPLETO con linaje. Incluye male + female + weather messages argentinos. Linaje (trazabilidad por sample) source: siempre "openslr61" subset: "main" (frases generales) o "weather" (mensajes de clima) gender: "m" / "f" speaker_id: ID anonimizado original del speaker file_id: ID original del archivo OpenSLR license: CC-BY-SA-4.0 Schema campo tipo… See the full description on the dataset page: https://huggingface.co/datasets/Kukedlc/openslr61-es-ar-full.audiotext-to-speech1K<n<10K0 likes52 downloads4mo agoHugging Face10fullmannger /atco2-asr-atcosim Dataset Card for "atco2-asr-atcosim" This is a dataset constructed from two datasets: ATCO2-ASR and ATCOSIM. It is divided into 80% train and 20% validation by selecting files randomly. Some of the files have additional information that is presented in the 'info' file. audioautomatic-speech-recognition10K<n<100K0 likes38 downloads3mo agoHugging Face11guizme /adlam_fulfulde Dataset Card for adlam fululde This dataset contains 51 Pulaar speech recordings. Credits and Acknowledgments This work was produced by the Organisation pour la promotion de la langue Pulaar (Winden jangen ADLaM). Contact Information Website: www.windenjangen.org Address: École Solokoure, Cimenterie, Conakry, Guinée Emails: * windenjangen@windenjangen.org aysha.sow12@gmail.com +224 622 15 40 75 +224 624463923 audioautomatic-speech-recognitionn<1K3 likes30 downloads10mo agoHugging Face12OcularAIInc /AMERICAN-ENGLISH-TRANSCRIBED-HIFI-FULL-DUPLEX-TWO-SPEAKER-CONVERSATIONAL-DATASET-SAMPLEgated AMERICAN ENGLISH TRANSCRIBED HI-FI FULL-DUPLEX TWO-SPEAKER CONVERSATIONAL DATASET — SAMPLE Overview This open sample from Ocular AI contains four American English conversations between two people, with a separate audio track for each speaker and verbatim transcripts containing segment- and word-level timestamps. The recordings capture conversational exchanges: repetitions, fillers, false starts, pauses, laughter, and audible breaths. Some conversations begin with… See the full description on the dataset page: https://huggingface.co/datasets/OcularAIInc/AMERICAN-ENGLISH-TRANSCRIBED-HIFI-FULL-DUPLEX-TWO-SPEAKER-CONVERSATIONAL-DATASET-SAMPLE.audioautomatic-speech-recognitionn<1K0 likes30 downloads4d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.