CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01aoxo /t2a-mommy t2a-mommy Female-voice ASMR corpus for the text2asmr project. Previously published as aoxo/audios2. Companion repos: aoxo/t2a-daddy (male voice), aoxo/t2a-audios-v1 (the original v1 corpus). Layout path what <creator>/<title>.m4a source audio, 48 kHz AAC, one folder per creator <creator>/<title>.json word-level Whisper large-v3 alignment ([] = skipped: near-silent or undecodable) labels/qwen3omni.jsonl non-speech ontology labels for gap clips… See the full description on the dataset page: https://huggingface.co/datasets/aoxo/t2a-mommy.audioaudio-classification0 likes15k downloads23m agoHugging Face02aoxo /t2a-daddy t2a-daddy Male-voice ASMR corpus for the text2asmr project. Previously published as aoxo/audios3. Companion repos: aoxo/t2a-mommy (female voice), aoxo/t2a-audios-v1 (the original v1 corpus). Layout path what <creator>/<title>.m4a source audio, 48 kHz AAC, one folder per creator <creator>/<title>.json word-level Whisper large-v3 alignment ([] = skipped: near-silent or undecodable) labels/qwen3omni.jsonl non-speech ontology labels for gap clips… See the full description on the dataset page: https://huggingface.co/datasets/aoxo/t2a-daddy.audioaudio-classification10K<n<100K0 likes3.2k downloads2h agoHugging Face03aoxo /t2a-audios-v1 t2a-audios-v1 The original text2asmr corpus (previously aoxo/audios): 48 kHz stereo ASMR audio with word-level alignments, used for the v1 generator (Chatterbox speech LoRA, Stable Audio Open trigger LoRA) and as the source for the reconstructed trigger ontology. Superseded for ontology work by aoxo/t2a-mommy and aoxo/t2a-daddy, which are larger, creator-attributed and split by voice. path what <id>.m4a source audio, 48 kHz <id>.json word-level alignment + silence… See the full description on the dataset page: https://huggingface.co/datasets/aoxo/t2a-audios-v1.audioaudio-classification0 likes1.8k downloads3d agoHugging Face04mteb /Urbansound8K_t2a Dataset Card for "Urbansound8K_t2a" More Information needed audio10K<n<100K0 likes951 downloads1y agoHugging Face05mteb /MACS_t2a Dataset Card for "MACS_t2a" More Information needed audio1K<n<10K0 likes867 downloads1y agoHugging Face06mteb /spoken-squad-t2aaudiotext-retrievaln<1K0 likes856 downloads8mo agoHugging Face07mteb /gigaspeech_t2aaudio10K<n<100K0 likes848 downloads1y agoHugging Face08glenn2 /legemma_t2t_data_544kaudio100K<n<1M0 likes283 downloads2y agoHugging Face09mteb /FLARE-1k-Unified-T2VAaudio1K<n<10K0 likes183 downloads2mo agoHugging Face10mteb /FLARE-1k-Audio-T2VAaudio1K<n<10K0 likes149 downloads2mo agoHugging Face11mteb /sounddescs_t2a SoundDescsT2ARetrieval An MTEB dataset Massive Text Embedding Benchmark Natural language description for different audio sources from the BBC Sound Effects webpage. Task category Any2AnyRetrieval (text-to-audio) Domains Encyclopaedic, Written Reference IEEE Transactions on Multimedia Source datasets: mteb/sounddescs_t2a How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task =… See the full description on the dataset page: https://huggingface.co/datasets/mteb/sounddescs_t2a.textother0 likes138 downloads5mo agoHugging Face12mteb /audiocaps_t2a AudioCapsT2ARetrieval An MTEB dataset Massive Text Embedding Benchmark Natural language description for any kind of audio in the wild. Task category t2a Domains Encyclopaedic, Written Reference https://audiocaps.github.io/ Source datasets: mteb/audiocaps_t2a How to evaluate on this task You can evaluate an embedding model on this dataset using the following code: import mteb task = mteb.get_task("AudioCapsT2ARetrieval") evaluator = mteb.MTEB([task])… See the full description on the dataset page: https://huggingface.co/datasets/mteb/audiocaps_t2a.audioother10K<n<100K0 likes134 downloads8mo agoHugging Face13Higobeatz /t2adata3audio10K<n<100K0 likes122 downloads2y agoHugging Face14lxercode /clotho_t2a_v2 ClothoT2ARetrieval.v2 An MTEB dataset Massive Text Embedding Benchmark An audio captioning dataset containing audio clips from the Freesound platform and their corresponding captions. Version 2 removes empty-string queries. For more information see #5062 Task category Any2AnyRetrieval (text-to-audio) Domains Encyclopaedic, Written Reference Clotho: An Audio Captioning Dataset Source datasets: mteb/Clotho mteb/Clotho How to evaluate on this task… See the full description on the dataset page: https://huggingface.co/datasets/lxercode/clotho_t2a_v2.audioother10K<n<100K0 likes101 downloads2mo agoHugging Face15glenn2 /legemma_t2t_data_544k_fullaudio100K<n<1M0 likes99 downloads2y agoHugging Face16aoxo /t2a-triggersaudio1K<n<10K0 likes61 downloads23h agoHugging Face17Gaie /t2a_audio_ldm2audio1K<n<10K0 likes56 downloads2y agoHugging Face18dukesun99 /SongDescriber-T2Aaudio1K<n<10K0 likes56 downloads2mo agoHugging Face19glenn2 /legemma_t2t_data_544k_full_trimaudio100K<n<1M0 likes55 downloads2y agoHugging Face20mteb /jl_corpus_t2aaudio1K<n<10K0 likes46 downloads1y agoHugging Face21mteb /MusicCaps_t2a Dataset Card for "MusicCaps_t2a" More Information needed audio10K<n<100K0 likes38 downloads1y agoHugging Face22hubxrt /LPMusicCapsMTT_t2a LPMusicCapsMTTT2ARetrieval An MTEB dataset Massive Text Embedding Benchmark LLM-generated pseudo captions for 10-second music clips from the MagnaTagATune dataset. Captions were produced by prompting a large language model with the human-annotated tags of each clip, giving four differently-styled captions per clip. Complements MusicCaps, whose captions are human-written and whose audio comes from AudioSet. Task category Any2AnyRetrieval (text-to-audio) Domains Music… See the full description on the dataset page: https://huggingface.co/datasets/hubxrt/LPMusicCapsMTT_t2a.audioother1K<n<10K0 likes35 downloads2mo agoHugging Face23mteb /LibriTTS_t2a Dataset Card for "LibriTTS_t2a" More Information needed audio10K<n<100K0 likes34 downloads1y agoHugging Face24Wissam42 /FLARE-1k-Unified-T2VAaudio1K<n<10K0 likes29 downloads2mo agoHugging Face25Wissam42 /FLARE-1k-Audio-T2VAaudio1K<n<10K0 likes29 downloads2mo agoHugging Face26Gaie /t2a_stable_audio_openaudio1K<n<10K0 likes28 downloads2y agoHugging Face2734data /v14-fake-vcapv-t2aaudio1K<n<10K0 likes26 downloads6mo agoHugging Face28mteb /EmoV_DB_t2a Dataset Card for "EmoV_DB_t2a" More Information needed audio1K<n<10K0 likes25 downloads1y agoHugging Face29Cybrpgs /corpus5-t2inserts-sample-200-20260916gatedaudioautomatic-speech-recognitionn<1K0 likes23 downloads10d agoHugging Face30arteemg /spoken-squad-t2aaudiotext-retrievaln<1K0 likes21 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.