CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ArlingtonCL2 /Barkopedia_Dog_Sex_Classification_Dataset 📦 Dataset Description This dataset is part of the Barkopedia Challenge: https://uta-acl2.github.io/barkopedia.html Check training data on Hugging Face: 👉 ArlingtonCL2/Barkopedia_Dog_Sex_Classification_Dataset This challenge provides a dataset of labeled dog bark audio clips: 29,345 total clips of vocalizations from 156 individual dogs across 5 breeds: Shiba Inu Husky Chihuahua German Shepherd Pitbull Training set: 26,895 clips 13,567 female13,328 male Test set: 2,450… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_Dog_Sex_Classification_Dataset.audioaudio-classification10K<n<100K0 likes5.2k downloads1y agoHugging Face02ArlingtonCL2 /Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Check Training Data here: ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Description This dataset is for Dog Age Group Classification and contains dog bark audio clips. The data is split into training, public test, and private test sets. Training set: 17888 audio clips. Test set: 4920 audio clips, further divided into: Test Public (~40%): 1966 audio clips for live leaderboard updates. Test Private (~60%): 2954 audio clips for final evaluation. You will… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K2 likes3.2k downloads1y agoHugging Face03vnahata /AfriMCQA-category-classification Afri-MCQA cross-modal cultural category classification (MTEB) Classify the cultural category of an entry from its photograph and the question about it spoken by a native speaker, across 16 African languages. Labels index this list: geography, building, and landmarks public figure and pop culture cooking and food objects, materials, clothing tranditions, art, and history brands, products, and companies plants and animals people, and everyday life vehicles and transportation… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/AfriMCQA-category-classification.audioaudio-classification1K<n<10K0 likes1.2k downloads19d agoHugging Face04hlx1021 /Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Check Training Data here: ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Description This dataset is for Dog Age Group Classification and contains dog bark audio clips. The data is split into training, public test, and private test sets. Training set: 17888 audio clips. Test set: 4920 audio clips, further divided into: Test Public (~40%): 1966 audio clips for live leaderboard updates. Test Private (~60%): 2954 audio clips for final evaluation. You… See the full description on the dataset page: https://huggingface.co/datasets/hlx1021/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes905 downloads3mo agoHugging Face05mteb /Vehicle_sounds_classification_datasetaudio1K<n<10K1 likes786 downloads8mo agoHugging Face06ArlingtonCL2 /Barkopedia_DOG_BREED_CLASSIFICATION_DATASET 📦 Dataset Description This dataset is part of the Barkopedia Challenge 🔗 https://uta-acl2.github.io/barkopedia.html Check Training Data here:👉 ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET This dataset contains 29,347 audio clips of dog barks labeled by dog breed. The audio samples come from 156 individual dogs across 5 dog breeds: shiba inu husky chihuahua german shepherd pitbull 📊 Per-Breed Summary Breed Train Public TestPrivate Test Test… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes679 downloads1y agoHugging Face07mesolitica /Zeroshot-Audio-Classification-Instructions Zeroshot-Audio-Classification-Instructions Convert audio classification dataset into zero-shot format speech instructions, support both single label and multi-label, VGGSound FSD50k Nonspeech7k urbansound8K VocalSound Emotion Gender ESD Emotion Age Language TAU Urban Acoustic Scenes 2022 CochlScene BirdCLEF_2021 EmoBox AudioSet We also converted huge WAV files into MP3 16k sample rate to reduce storage size.To prevent leakage, please do not include test set in training session.… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Zeroshot-Audio-Classification-Instructions.audio1M<n<10M3 likes554 downloads1y agoHugging Face08rpmon /fma-genre-classification FMA Genre Classification Dataset The FMA Genre Classification Dataset is a subset of the Free Music Archive (FMA), containing audio samples and genre labels for music classification tasks. This version uses the "small" subset of FMA, which contains 8,000 tracks of 30 seconds each, evenly distributed across 8 genres. Dataset Description Dataset Summary This dataset consists of 8,000 audio tracks from the Free Music Archive (FMA), each 30 seconds in length… See the full description on the dataset page: https://huggingface.co/datasets/rpmon/fma-genre-classification.audio1K<n<10K3 likes441 downloads2y agoHugging Face09mesolitica /Classification-Speech-Instructions Classification Speech Instructions Speech instructions for emotion, gender, age and language audio classification. Source code Source code at https://github.com/mesolitica/malaysian-dataset/tree/master/llm-instruction/speech-classification-instructions audioaudio-classification100K<n<1M1 likes269 downloads1y agoHugging Face10PuneettArora /Barkopedia_DOG_BREED_CLASSIFICATION_DATASET 📦 Dataset Description This dataset is part of the Barkopedia Challenge 🔗 https://uta-acl2.github.io/barkopedia.html Check Training Data here:👉 ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET This dataset contains 29,347 audio clips of dog barks labeled by dog breed. The audio samples come from 156 individual dogs across 5 dog breeds: shiba inu husky chihuahua german shepherd pitbull 📊 Per-Breed Summary Breed Train Public Test Private Test… See the full description on the dataset page: https://huggingface.co/datasets/PuneettArora/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes253 downloads16d agoHugging Face11hamsaai /Recorrected_Classification_Data_filtered_trainaudio10K<n<100K0 likes238 downloads2mo agoHugging Face12vnahata /CAMEO-emotion-classification CAMEO multilingual speech emotion classification (MTEB) Speech emotion recognition across five languages, drawn from the CAMEO collection. Labels index this list: anger fear happiness neutral sadness surprise Source: amu-cai/CAMEO at revision 38e9e96, cc-by-nc-sa-4.0. Split by speaker so no speaker appears in both train and test. Only the six emotions common to every included language are kept. Audio is 16 kHz Opus. Built by scripts/data/cameo_emotion/create_data.py in the… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/CAMEO-emotion-classification.audioaudio-classification10K<n<100K0 likes140 downloads20d agoHugging Face13hamsaai /Recorrected_Classification_Data_filtered_train_22audio10K<n<100K0 likes138 downloads2mo agoHugging Face14kudukudu /building_floor_classificationDataset_chunked_5 : chunks of 05 seconds obtained from expert samples Dataset_chunked_10 : chunks of 10 seconds obtained from expert samples Dataset_expanded : chunks of 10 seconds obtained from whole samples Data.zip : original dataset audio1K<n<10K0 likes130 downloads4y agoHugging Face15vhands /audio-event-classification-post-public audio-event-classification-post-public Sound-event and acoustic-scene classification annotations: ESC-50 (environmental), UrbanSound8K, FSD50k (50k+ events), TUT-Acoustic-Scenes-2017, DCASE-2025, NonSpeech7k (vocal sounds), VocalSound (laugh/cough/sigh). Useful for training audio LLMs on the perception substrate underneath higher-level reasoning. Audio is not bundled in this repo. See download.sh and per-dataset data/<name>.info.json for the fetch recipe; run postlink_audio.py… See the full description on the dataset page: https://huggingface.co/datasets/vhands/audio-event-classification-post-public.textaudio-classification100K<n<1M1 likes130 downloads3mo agoHugging Face16laion /vocal-burst-classification-v2 Vocal Burst Classification V2 — laion/vocal-burst-classification-v2 The V2 training corpus for the Vocal Burst Classifier V2: a single-label dataset over an 83-class vocal-burst taxonomy (82 non-speech human vocalizations no_burst, index 82), shipped as precomputed VoiceCLAP-commercial embeddings plus the raw vocal-bursts-clean audio. The vocal-burst clips were generated with various synthetic text-to-audio models such as DramaBox, then annotated and filtered as described… See the full description on the dataset page: https://huggingface.co/datasets/laion/vocal-burst-classification-v2.audioaudio-classification10K<n<100K0 likes113 downloads2mo agoHugging Face17MUGEN-Benchmark /Genre_Classificationaudion<1K0 likes111 downloads8mo agoHugging Face18DynamicSuperb /Vehicle_sounds_classification_datasetaudio1K<n<10K1 likes89 downloads2y agoHugging Face19danilotpnta /GTZAN_genre_classificationaudion<1K0 likes75 downloads2y agoHugging Face20OpenWhistleNeurIPS26 /OpenWhistle-Classification-Finetuning OpenWhistle Classification Finetuning Dataset OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning is the public classification finetuning dataset used for dolphin whistle identity classification. It contains short whistle clips, whistle-level metadata, fundamental-frequency tracks, rendered F0 spectrograms, and integer class labels. The main reviewer-facing subset is the balanced balanced config. It contains six classes: NSW_1 (label=0) SW_Luna (label=1) SW_Nana (label=2)… See the full description on the dataset page: https://huggingface.co/datasets/OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning.audioaudio-classification10K<n<100K0 likes74 downloads5mo agoHugging Face21hr16 /ViSpeech-Gender-Dialect-Classificationimport datasets as hugDS import pandas as pd import os os.environ["HF_HUB_ENABLE_HF_TRANSFER"] = "1" from df.io import resample from df.enhance import enhance, init_df import torch import warnings df_model, df_state, _ = init_df() SAMPLING_RATE = 16_000 def normalize_vietmed(example): global vietmed_info example["gender"] = vietmed_info[vietmed_info["Speaker ID"] == example["Speaker ID"]]["Gender"].values[0].lower() example["dialect"] = vietmed_info[vietmed_info["Speaker ID"] ==… See the full description on the dataset page: https://huggingface.co/datasets/hr16/ViSpeech-Gender-Dialect-Classification.audioaudio-classification10K<n<100K2 likes71 downloads2y agoHugging Face22danki2meme /Audio_for_age_classification_Evalaudion<1K1 likes71 downloads7mo agoHugging Face23vocsim /mouse-identity-classification-benchmark VocSim — Mouse Identity Classification A companion dataset for the VocSim benchmark that tests whether audio embeddings preserve individual identity in mouse ultrasonic vocalizations (USVs). It contains pre-segmented USV syllables from multiple individual mice (the speaker field), sampled at the native 250 kHz, derived from recordings by Van Segbroeck et al. (2017). Basha, M., Zai, A. T., Stoll, S., & Hahnloser, R. H. R. VocSim: A Training-free Benchmark for Zero-shot Content… See the full description on the dataset page: https://huggingface.co/datasets/vocsim/mouse-identity-classification-benchmark.audio10K<n<100K0 likes69 downloads4mo agoHugging Face24danki2meme /Audio_for_age_classification_Trainaudio1K<n<10K4 likes69 downloads7mo agoHugging Face25greenarcade /wav2vec2-vd-bird-sound-classification-datasetaudioaudio-classification1K<n<10K3 likes56 downloads1y agoHugging Face26meghana-007 /tamil-audio-emotion-classificationaudioaudio-classificationn<1K0 likes49 downloads6mo agoHugging Face27Arulpandi /audio_classification_dataset2audion<1K1 likes48 downloads1y agoHugging Face28MUGEN-Benchmark /Gender_Classificationaudion<1K1 likes47 downloads8mo agoHugging Face29kuross /dl-proj-classificationaudio1K<n<10K0 likes45 downloads10mo agoHugging Face30hamsaai /Recorrected_Classification_Data_filtered_syr_validatedaudio1K<n<10K0 likes43 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.