CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MVRL /GeoSound GeoSound GeoSound is a geo-referenced soundscape dataset that pairs satellite/aerial imagery with environmental audio recordings. It aggregates recordings from four crowdsourcing platforms — Freesound, Aporee, iNaturalist, and Flickr (via the YFCC100M collection) — and covers a wide geographic footprint. Research use only. See LICENSE.md for full license and attribution details. Splits Split Rows train293,718 val 4,999 test 9,931 Train/val/test… See the full description on the dataset page: https://huggingface.co/datasets/MVRL/GeoSound.audioaudio-classification100K<n<1M1 likes1.2k downloads5mo agoHugging Face02kojo-george /asante-twi-ttsaudio10K<n<100K0 likes404 downloads2y agoHugging Face03kojo-george /akuapem-twi-ttsaudio10K<n<100K1 likes312 downloads2y agoHugging Face04georgechang8 /code_switch_yodas_zh Dataset Card for code-switching yodas This dataset is derived from espnet/yodas, more details can be found here: https://huggingface.co/datasets/espnet/yodas This is a subset of the zh000 subset of espnet/yodas dataset, which selects videos with Mandarin-English code-switching phenomenon. Note that code-switching is only gauranteed per video rather than per utterance. Therefore, not every utterance in the dataset contains code-switching. Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/georgechang8/code_switch_yodas_zh.audio10K<n<100K4 likes185 downloads2y agoHugging Face05NMikka /Common-Voice-Geo-Cleaned Common Voice Georgian — Cleaned for TTS/STT A high-quality subset of Mozilla Common Voice Georgian cleaned and filtered specifically for text-to-speech fine-tuning. Dataset Summary Total samples 21,421 Total duration 35.0 hours Speakers 12 Sample rate 24 kHz mono WAV Language Georgian (kat) Source Mozilla Common Voice 19.0 License CC-0 (public domain) Splits Split Samples Description train 20,300 Training data eval… See the full description on the dataset page: https://huggingface.co/datasets/NMikka/Common-Voice-Geo-Cleaned.audiotext-to-speech10K<n<100K9 likes167 downloads7mo agoHugging Face06geoffbremneraudio /Geoff_Bremner_Multimodal_Music_Corpus_SAMPLE Geoff Bremner Multimodal Music Corpus — Sample Release This is a single-track sample from the Geoff Bremner Multimodal Music Corpus, a growing, research-grade, commercially licensable dataset of 100% original music — written, recorded, and produced entirely by one artist . If this sample meets your needs - please contact me directly for more Geoff Bremner https://linktr.ee/gbaudio License This dataset is released under CC BY-NC 4.0… See the full description on the dataset page: https://huggingface.co/datasets/geoffbremneraudio/Geoff_Bremner_Multimodal_Music_Corpus_SAMPLE.audion<1K2 likes119 downloads3mo agoHugging Face07CH0BRK /Common-Voice-Geo-Cleaned Common Voice Georgian — Cleaned for TTS/STT A high-quality subset of Mozilla Common Voice Georgian cleaned and filtered specifically for text-to-speech fine-tuning. Dataset Summary Total samples 21,421 Total duration 35.0 hours Speakers 12 Sample rate 24 kHz mono WAV Language Georgian (kat) Source Mozilla Common Voice 19.0 License CC-0 (public domain) Splits Split Samples Description train 20,300 Training… See the full description on the dataset page: https://huggingface.co/datasets/CH0BRK/Common-Voice-Geo-Cleaned.audiotext-to-speech10K<n<100K0 likes82 downloads3mo agoHugging Face08georgechang8 /ASCEND_CLEAN Dataset Card for Dataset Name This dataset is derived from CAiRE/ASCEND. More information is available at https://huggingface.co/datasets/CAiRE/ASCEND. Removed 嗯 呃 um uh Resolved [UNK]'s using whisper-medium Usage Default utterances with cleaned transcripts from datasets import load_dataset data = load_dataset("georgechang8/ASCEND_CLEAN") # add split="train" for train set, etc. Concatenated 30s utterances with cleaned transcripts… See the full description on the dataset page: https://huggingface.co/datasets/georgechang8/ASCEND_CLEAN.audio10K<n<100K0 likes76 downloads2y agoHugging Face09georgechang8 /cv16_30saudio1K<n<10K0 likes48 downloads2y agoHugging Face10Georgejohn /GTZAN Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/Georgejohn/GTZAN.audio1K<n<10K0 likes30 downloads2y agoHugging Face11norjas1 /GEOdatasetaudio0 likes27 downloads2y agoHugging Face12Speech-data /Georgian-Speech-Dataset Field Value 📜 License CC BY-NC-ND 4.0 🎯 Task Categories Automatic Speech Recognition 🌍 Language Georgian (ka) 🏷️ Tags Audio, Speech, Speech Recognition, Georgian, ML, Machine, Machine Learning 📦 Size Category n < 1K audioautomatic-speech-recognitionn<1K0 likes20 downloads6mo agoHugging Face13shunyalabs /georgian-speech-datasetaudio1K<n<10K1 likes15 downloads1y agoHugging Face14akalandia /ka-geo-voice-male-v1 Dataset Card for Georgian Male Voice Dataset v1 Intended Use Primary Use: Training and fine-tuning TTS models for Georgian language synthesis, including microsoft/speecht5_tts. Secondary Use: Research in speech synthesis, voice conversion, or linguistic analysis. SpeechT5 Compatibility This dataset is specifically formatted to be compatible with microsoft/speecht5_tts fine-tuning. The dataset includes: Audio: 22,050 Hz mono WAV files (matching SpeechT5… See the full description on the dataset page: https://huggingface.co/datasets/akalandia/ka-geo-voice-male-v1.audiotext-to-speechn<1K0 likes9 downloads9mo agoHugging Face15GeoPoll /dataset-20250728_102101-swgated GeoPoll Swahili Speech Dataset This dataset contains speech recognition data for Swahili (sw) collected and processed by GeoPoll. Dataset Summary This dataset is designed for fine-tuning speech recognition models on Swahili audio data. It includes high-quality audio segments with corresponding transcriptions. Dataset Statistics Total samples: 11814 Total duration: 20.45 hours Average duration: 6.23 seconds per sample Number of speakers: 6 Language: Swahili… See the full description on the dataset page: https://huggingface.co/datasets/GeoPoll/dataset-20250728_102101-sw.audioautomatic-speech-recognition10K<n<100K0 likes7 downloads1y agoHugging Face16geojacob /Testaudion<1K0 likes7 downloads1y agoHugging Face17velocity-engg /Style_TTS_UPSC_GEOaudio1K<n<10K0 likes6 downloads2y agoHugging Face18Geofab /ArtieAbamsaudion<1K0 likes4 downloads3y agoHugging Face19geovanezzz /vozviniboyaudion<1K0 likes4 downloads3y agoHugging Face20Manish2649 /TTS_10s_clean_documentry_style_national_geographyaudion<1K0 likes3 downloads5mo agoHugging Face21Albinator /George-Harrison-Brainwashed-TheOsloChildaudion<1K0 likes3 downloads4mo agoHugging Face22archivartaunik /georgii-marchuk-davyd-garadotskiia-kanony Давыд-Гарадоцкія каноны Metadata Author: Георгій Марчук Title: Давыд-Гарадоцкія каноны Narrator: Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/georgii-marchuk-davyd-garadotskiia-kanony.audion<1K0 likes2 downloads4mo agoHugging Face23archivartaunik /georgii-marchuk-kryk-na-khutary-margaryta-zakharyia Крык на хутары Metadata Author: Георгій Марчук Title: Крык на хутары Narrator: Маргарыта Захарыя Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/georgii-marchuk-kryk-na-khutary-margaryta-zakharyia.audion<1K0 likes2 downloads4mo agoHugging Face24Albinator /GeorgeHarrisonTalkingaudion<1K0 likes1 downloads1y agoHugging Face25GLauzza /Mixed_Training_Geogatedaudion<1K0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.