CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Yi3852 /midi-audio-abc_300smidi, synthesized audio, ABC code triples (this dataset contains those with audio duration in 5-300s, several subsets with smaller duration 60s 30s 10s) (token_length_abc field represents the token count of the abc text w.r.t. Qwen3's tokenizer) midi files are from bread-midi-dataset synthesized audio: use Don Allen's Timbres of Heaven as soundfont and FluidSynth as synthesizer abc notation: mid2abc by EasyABC (midi2abc.py) Citation @misc{jiang2025advancingfoundationmodelmusic… See the full description on the dataset page: https://huggingface.co/datasets/Yi3852/midi-audio-abc_300s.audio100K<n<1M0 likes463 downloads1y agoHugging Face02Yi3852 /midi-audio-abc_60smidi, synthesized audio, ABC code triples (this dataset contains those with audio duration in 5-60s, sampled from the full set with max 300s duration) (token_length_abc field represents the token count of the abc text w.r.t. Qwen3's tokenizer) midi files are from bread-midi-dataset synthesized audio: use Don Allen's Timbres of Heaven as soundfont and FluidSynth as synthesizer abc notation: mid2abc by EasyABC (midi2abc.py) Citation @misc{jiang2025advancingfoundationmodelmusic… See the full description on the dataset page: https://huggingface.co/datasets/Yi3852/midi-audio-abc_60s.audio100K<n<1M0 likes413 downloads1y agoHugging Face03Yi3852 /midi-audio-abc_longmidi, synthesized audio, ABC code triples (this dataset contains those with audio duration in 5 min - 2 hours, less than 5 min data are in 300s and there are several subsets with smaller duration 60s 30s 10s) (token_length_abc field represents the token count of the abc text w.r.t. Qwen3's tokenizer) midi files are from bread-midi-dataset synthesized audio: use Don Allen's Timbres of Heaven as soundfont and FluidSynth as synthesizer abc notation: mid2abc by EasyABC (midi2abc.py)… See the full description on the dataset page: https://huggingface.co/datasets/Yi3852/midi-audio-abc_long.audio10K<n<100K1 likes389 downloads1y agoHugging Face04Yi3852 /midi-audio-abc_30smidi, synthesized audio, ABC code triples (this dataset contains those with audio duration in 5-30s, sampled from the full set with max 300s duration) (token_length_abc field represents the token count of the abc text w.r.t. Qwen3's tokenizer) midi files are from bread-midi-dataset synthesized audio: use Don Allen's Timbres of Heaven as soundfont and FluidSynth as synthesizer abc notation: mid2abc by EasyABC (midi2abc.py) Citation @misc{jiang2025advancingfoundationmodelmusic… See the full description on the dataset page: https://huggingface.co/datasets/Yi3852/midi-audio-abc_30s.audio10K<n<100K0 likes136 downloads1y agoHugging Face05Abcdefghijklmnopqrstuvwxyz12 /MODELOSDETESTEaudio0 likes96 downloads1y agoHugging Face06abc-123-456 /minds14 MInDS-14 MINDS-14 is training and evaluation resource for intent detection task with spoken data. It covers 14 intents extracted from a commercial system in the e-banking domain, associated with spoken examples in 14 diverse language varieties. Example MInDS-14 can be downloaded and used as follows: from datasets import load_dataset minds_14 = load_dataset("PolyAI/minds14", "fr-FR") # for French # to download all data for multi-lingual fine-tuning uncomment… See the full description on the dataset page: https://huggingface.co/datasets/abc-123-456/minds14.audioautomatic-speech-recognition10K<n<100K0 likes78 downloads19d agoHugging Face07Yi3852 /midi-audio-abc_10smidi, synthesized audio, ABC code triples (this dataset contains those with audio duration in 5-10s, sampled from the full set with max 300s duration) (token_length_abc field represents the token count of the abc text w.r.t. Qwen3's tokenizer) midi files are from bread-midi-dataset synthesized audio: use Don Allen's Timbres of Heaven as soundfont and FluidSynth as synthesizer abc notation: mid2abc by EasyABC (midi2abc.py) Citation @misc{jiang2025advancingfoundationmodelmusic… See the full description on the dataset page: https://huggingface.co/datasets/Yi3852/midi-audio-abc_10s.audio10K<n<100K0 likes25 downloads1y agoHugging Face08hadramibegnoug /abc.hassaniya_ASRaudio1K<n<10K0 likes14 downloads1y agoHugging Face09learn-abc /indian-speech-audio-extendedaudion<1K0 likes13 downloads1y agoHugging Face10sw-voice /swamiji-artifact-abc-grid Where does the end-of-clip artifact come from? Reported symptom: short polite replies "mess up at the end", and a "huge high" is audible after the words finish — sometimes even when the sentence ends in a full stop. Three candidate causes: the model, the streaming, or the Opus codec. Each phrase here is generated twice — ending in ! and ending in . — and each generation is rendered three ways, so exactly one variable moves at a time. Six players per row. rendering what it… See the full description on the dataset page: https://huggingface.co/datasets/sw-voice/swamiji-artifact-abc-grid.audiotext-to-speechn<1K0 likes10 downloads2mo agoHugging Face11Abcdefghijklmnopqrstuvwxyz12 /BABYMONSTERTESTEaudion<1K0 likes9 downloads2y agoHugging Face12a-b-c-nick /vocalsound-throat-sneezeaudio1K<n<10K0 likes7 downloads1y agoHugging Face13continueawj /abcaudion<1K0 likes7 downloads1y agoHugging Face14Abcdefghijklmnopqrstuvwxyz12 /MODELSTESTE2audion<1K0 likes5 downloads1y agoHugging Face15verbreb /abcs_16k_headset_templeaudio10K<n<100K0 likes3 downloads4mo agoHugging Face16verbreb /vibravox_abcs_merge_headset_temple_16kaudio10K<n<100K0 likes3 downloads4mo agoHugging Face17idrakai1 /abc_datagatedaudio100K<n<1M1 likes2 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.