CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Aalto-Speech-Synthesis /icelandic_asr Icelandic ASR Collection This repository collects six Icelandic speech corpora in directly loadable Parquet form. Audio is embedded as 16 kHz mono FLAC bytes. The repository is a convenience repackaging: the linked CLARIN-IS records and original dataset repositories remain the canonical sources and should be cited when using the data. No configuration is selected by default. Choose a corpus configuration and, for this large collection, normally choose a split explicitly.… See the full description on the dataset page: https://huggingface.co/datasets/Aalto-Speech-Synthesis/icelandic_asr.audioautomatic-speech-recognition1M<n<10M0 likes507 downloads19d agoHugging Face02Aviv-anthonnyolime /SIWIS_French_Speech_Synthesis_Database SIWIS French Speech Synthesis Database This README provides a concise description of the dataset, including its structure, file naming conventions, and known labeling issues. Additionally, suggestions for potential improvements are outlined in the TODO section. The dataset is distributed under the Creative Commons Attribution 4.0 International (CC BY 4.0) license, permitting its use for any purpose. For more details about the database design and recording process, please refer… See the full description on the dataset page: https://huggingface.co/datasets/Aviv-anthonnyolime/SIWIS_French_Speech_Synthesis_Database.audioautomatic-speech-recognition10K<n<100K0 likes301 downloads2y agoHugging Face03jsun39 /enni-child-speech-synthesislicense: mit task_categories: text-to-speech automatic-speech-recognition language: en tags: speech audio child-speech talkbank size_categories: 10K<n<100K TalkBank Child Speech Synthesis Dataset (Seed 1) This dataset contains child speech synthesis data generated from the TalkBank FASA ENNI corpus. Dataset Information Number of Samples: 10032 Seed: 1 Audio Format: WAV (16kHz) Source: TalkBank FASA ENNI Data Structure The dataset contains the following… See the full description on the dataset page: https://huggingface.co/datasets/jsun39/enni-child-speech-synthesis.audio1K<n<10K7 likes221 downloads6mo agoHugging Face04Aalto-Speech-Synthesis /stortinget_speech_corpus_v1.0 Dataset Card for Stortinget Speech Corpus V1.0 Overview This is the WebDataset version of the Stortinget Speech Corpus V1.0, originally created by the National Library of Norway. We re-organize it into WebDataset format for better usability. The Stortinget Speech Corpus (SSC) is a 5000+ hours speech dataset for weak supervision ASR created from audio andaligned proceedings text from Stortinget, the Norwegian Parliament. For more information, please refer to the original… See the full description on the dataset page: https://huggingface.co/datasets/Aalto-Speech-Synthesis/stortinget_speech_corpus_v1.0.audioautomatic-speech-recognition100K<n<1M0 likes83 downloads5mo agoHugging Face05VladS159 /common_voice_16_1_romanian_speech_synthesisaudio10K<n<100K0 likes47 downloads3y agoHugging Face06VladS159 /common_voice_17_0_romanian_speech_synthesisaudio10K<n<100K1 likes41 downloads2y agoHugging Face07VladS159 /common_voice_romanian_speech_synthesisaudio10K<n<100K2 likes40 downloads3y agoHugging Face08doannv /speech-synthesisgatedaudio100K<n<1M0 likes24 downloads6mo agoHugging Face09doannv /tts-synthesis-audiogatedaudio100K<n<1M0 likes3 downloads7mo agoHugging Face10CAS-SIAT-XinHai /AudiPsy-Synthesisgated AudiPsy-Synthesis: Multilingual Emotional Counseling Dialogue Dataset AudiPsy is a multilingual emotional counseling dialogue dataset containing paired speech audio and text transcripts for psychological counseling conversations. The dataset includes synthetic speech generated using modern TTS systems and annotated emotional information. It is designed to support research in: emotional speech understanding mental health dialogue modeling speech-text multimodal learning counseling… See the full description on the dataset page: https://huggingface.co/datasets/CAS-SIAT-XinHai/AudiPsy-Synthesis.audioautomatic-speech-recognition10K<n<100K0 likes2 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.