CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01michsethowusu /vai-speech-text-parallel Vai Speech-Text Parallel Dataset Dataset Description This dataset contains 23286 parallel speech-text pairs for Vai, a language spoken primarily in Ghana. The dataset consists of audio recordings paired with their corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Dataset Summary Language: Vai - vai Task: Speech Recognition, Text-to-Speech Size: 23286 audio files > 1KB (small/corrupted… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/vai-speech-text-parallel.audioautomatic-speech-recognition10K<n<100K0 likes223 downloads1y agoHugging Face02doof-ferb /vais1000 unofficial mirror of VAIS-1000 official announcement: https://vais.vn/vi/tai-ve/hts_for_vietnamese (dead) mirror: https://github.com/undertheseanlp/text_to_speech/tree/run/data/vais1000/raw small only 1h40min audio - 1 speaker (female northern accent) - 1k samples pre-process: none need to do: check misspelling, restore foreign words phonetised to vietnamese usage with HuggingFace: # pip install -q "datasets[audio]" from datasets import load_dataset from torch.utils.data import… See the full description on the dataset page: https://huggingface.co/datasets/doof-ferb/vais1000.audioautomatic-speech-recognition1K<n<10K0 likes54 downloads2y agoHugging Face03vaishnavikedar4 /MCIF Dataset Description, Collection, and Source MCIF (Multimodal Crosslingual Instruction Following) is a multilingual human-annotated benchmark based on scientific talks that is designed to evaluate instruction-following in crosslingual, multimodal settings over both short- and long-form inputs. MCIF spans three core modalities -- speech, vision, and text -- and four diverse languages (English, German, Italian, and Chinese), enabling a comprehensive evaluation of MLLMs'… See the full description on the dataset page: https://huggingface.co/datasets/vaishnavikedar4/MCIF.audioautomatic-speech-recognition1K<n<10K0 likes39 downloads9mo agoHugging Face04vaidaryan13 /story_audio_4hr_15_diaaudion<1K0 likes21 downloads1y agoHugging Face05abar-uwc /vaani-bihar_vaishali-cleanedaudio1K<n<10K0 likes21 downloads1y agoHugging Face06vaidaryan13 /story_audio_4hr_15secaudion<1K1 likes16 downloads1y agoHugging Face07vaidaryan13 /dia_american_male_10audio1K<n<10K0 likes13 downloads1y agoHugging Face08vaitom /toynowaudion<1K0 likes12 downloads3y agoHugging Face09Vaidik7781 /sarvam-tts-dataset Sarvam TTS Training Dataset High-quality TTS training dataset built for expressive speech synthesis. Stats Total: 377 segments | 167.9 minutes English (en-IN): 191 segments | 84.3 minutes Hindi (hi-IN): 186 segments | 83.6 minutes Rejection rate: 26.9% after full manual human review of all 483 segments Emotion Distribution neutral: 259 | calm: 36 | sad: 32 | angry: 23 | happy: 12 | surprised: 6 | fearful: 6 | excited: 3 How it was… See the full description on the dataset page: https://huggingface.co/datasets/Vaidik7781/sarvam-tts-dataset.audion<1K0 likes10 downloads4mo agoHugging Face10archivartaunik /mikhas-lynkou-pra-smelaga-vaiaku-mishku-i-iago-slaunykh-tavaryshau Пра смелага ваяку Мішку і яго слаўных таварышаў Metadata Author: Міхась Лынькоў Title: Пра смелага ваяку Мішку і яго слаўных таварышаў Narrator: Source Group: Дзіцячыя Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/mikhas-lynkou-pra-smelaga-vaiaku-mishku-i-iago-slaunykh-tavaryshau.audion<1K0 likes10 downloads4mo agoHugging Face11vaidaryan13 /indic_hindi_new_2audion<1K0 likes9 downloads1y agoHugging Face12DZN111 /vaiaudion<1K0 likes8 downloads3y agoHugging Face13vaidaryan13 /story_audio_4hr_15_dia_10secaudio1K<n<10K0 likes8 downloads1y agoHugging Face14vaidaryan13 /dia_british_female_10audio1K<n<10K0 likes8 downloads1y agoHugging Face15alexpanick /vaiNeymaraudion<1K0 likes7 downloads3y agoHugging Face16vaidaryan13 /jaishankaraudion<1K0 likes7 downloads1y agoHugging Face17vaidaryan13 /dia_american_female_10audio1K<n<10K0 likes6 downloads1y agoHugging Face18vaishnavvv /Baby-Cry-Classification-Baiduaudio1K<n<10K0 likes6 downloads7mo agoHugging Face19archivartaunik /maksim-garetski-na-imperyalistychnai-vaine-maksim-viniarski На імперыалістычнай вайне Metadata Author: Максім Гарэцкі Title: На імперыалістычнай вайне Narrator: Максім Вінярскі Source Group: Аўдыёкнігі Source: rutracker.org Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/maksim-garetski-na-imperyalistychnai-vaine-maksim-viniarski.audion<1K0 likes6 downloads4mo agoHugging Face20archivartaunik /ivan-navumenka-khloptsy-samai-vialikai-vainy-andrei-kaliada Хлопцы самай вялікай вайны Metadata Author: Іван Навуменка Title: Хлопцы самай вялікай вайны Narrator: Андрэй Каляда Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-navumenka-khloptsy-samai-vialikai-vainy-andrei-kaliada.audion<1K0 likes6 downloads4mo agoHugging Face21vaidaryan13 /my_audio_datasetaudion<1K0 likes5 downloads1y agoHugging Face22vaidaryan13 /indic_hindiaudion<1K1 likes5 downloads1y agoHugging Face23vaidaryan13 /indic_hindi_finalaudion<1K0 likes4 downloads1y agoHugging Face24vaidaryan13 /dia_british_10audio1K<n<10K0 likes4 downloads1y agoHugging Face25Kontact /vais1000_sidon_noise_removalaudio1K<n<10K0 likes4 downloads5mo agoHugging Face26VAISHNAWI1 /vaishnawi_tts_cleaned_audioaudion<1K0 likes3 downloads1y agoHugging Face27archivartaunik /eva-vezhnavets-pa-shto-idzesh-voucha-zui-vaitsiakhouskaia Па што ідзеш, воўча Metadata Author: Ева Вежнавец Title: Па што ідзеш, воўча Narrator: Зуй-Вайцяхоўская Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/eva-vezhnavets-pa-shto-idzesh-voucha-zui-vaitsiakhouskaia.audion<1K0 likes3 downloads4mo agoHugging Face28archivartaunik /iuia-vislander-tumas-vislander-zmitser-vaitsiushkevich Тумас Вісландэр Metadata Author: Юя Вісландэр Title: Тумас Вісландэр Narrator: Зміцер Вайцюшкевіч Source Group: Дзіцячыя Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/iuia-vislander-tumas-vislander-zmitser-vaitsiushkevich.audio0 likes2 downloads4mo agoHugging Face29archivartaunik /mikhas-lynkou-pra-smelaga-vaiaku-mishku Пра сьмелага ваяку Мішку Metadata Author: Міхась Лынькоў Title: Пра сьмелага ваяку Мішку Narrator: Source Group: Дзіцячыя Source: http://staroeradio.ru Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/mikhas-lynkou-pra-smelaga-vaiaku-mishku.audion<1K0 likes2 downloads4mo agoHugging Face30archivartaunik /aliaksandar-vaitovich-mae-akademii Мае Акадэміі Metadata Author: Аляксандар Вайтовіч Title: Мае Акадэміі Narrator: Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size: about 250 MB.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/aliaksandar-vaitovich-mae-akademii.audion<1K0 likes2 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.