CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nccratliri /vad-animals Positive Transfer Of The Whisper Speech Transformer To Human And Animal Voice Activity Detection We proposed WhisperSeg, utilizing the Whisper Transformer pre-trained for Automatic Speech Recognition (ASR) for both human and animal Voice Activity Detection (VAD). For more details, please refer to our paper Positive Transfer of the Whisper Speech Transformer to Human and Animal Voice Activity Detection Nianlong Gu, Kanghwi Lee, Maris Basha, Sumit Kumar Ram, Guanghao You, Richard… See the full description on the dataset page: https://huggingface.co/datasets/nccratliri/vad-animals.audio0 likes3.7k downloads3y agoHugging Face02cgeorgiaw /animal-sounds Animal Sounds Collection This dataset contains audio recordings of various animal vocalizations from a range of species, curated to support research in bioacoustics, species classification, and sound event detection. It includes clean and annotated audio samples from the following animals: Birds Dogs Egyptian fruit bats Giant otters Macaques Orcas Zebra finches The dataset is designed to be lightweight and modular, making it easy to explore cross-species vocal… See the full description on the dataset page: https://huggingface.co/datasets/cgeorgiaw/animal-sounds.audio10K<n<100K20 likes3.4k downloads1y agoHugging Face03Hariprasath5128 /marine-animals-multimodal-dataset Marine Animals Multimodal Dataset 🐋 A comprehensive multimodal dataset combining audio recordings and images of 32 marine species. Dataset Summary Total samples: 24,911 Species: 32 Audio files: 1,357 unique recordings Images: 581 (309 matched + 272 from iNaturalist) Features species (string): Species name label (int32): Numeric label (0–31) audio (Audio): Audio recording of the species image (Image): Species image image_index (int32): Image number… See the full description on the dataset page: https://huggingface.co/datasets/Hariprasath5128/marine-animals-multimodal-dataset.audioaudio-classification10K<n<100K0 likes634 downloads9mo agoHugging Face04hdong51 /Human-Animal-Cartoon Human-Animal-Cartoon dataset Our Human-Animal-Cartoon (HAC) dataset consists of seven actions (‘sleeping’, ‘watching tv’, ‘eating’, ‘drinking’, ‘swimming’, ‘running’, and ‘opening door’) performed by humans, animals, and cartoon figures, forming three different domains. We collect 3381 video clips from the internet with around 1000 for each domain and provide three modalities in our dataset: video, audio, and pre-computed optical flow. The dataset can be used for Multi-modal Domain… See the full description on the dataset page: https://huggingface.co/datasets/hdong51/Human-Animal-Cartoon.audiozero-shot-classification1K<n<10K5 likes544 downloads5mo agoHugging Face05mesolitica /Animal-Sound-Instructions Animal Sound Instructions We gathered from, Birds, birdclef-2021 Insecta, christopher/birdclef-2025 Amphibia, christopher/birdclef-2025 Mammalia, christopher/birdclef-2025 We use Qwen/Qwen2.5-72B-Instruct to generate the answers based on the metadata. how to prepare the dataset huggingface-cli download \ mesolitica/Animal-Sound-Instructions \ --include "*.zip" \ --repo-type "dataset" \ --local-dir './' wget… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Animal-Sound-Instructions.audio10K<n<100K0 likes225 downloads1y agoHugging Face06monish-73 /marine-animals-multimodalaudio10K<n<100K0 likes106 downloads10mo agoHugging Face07DynamicSuperb /EnvironmentalSoundClassification_ESC50-Animals Dataset Card for "environmental_sound_classification_animals_ESC50" More Information needed audion<1K1 likes88 downloads3y agoHugging Face08Tri-PvP /animalaudion<1K0 likes84 downloads5mo agoHugging Face09enyoukai /AudioSet-Strong-Animalsaudio10K<n<100K0 likes83 downloads1y agoHugging Face10zachz /Human-Animal-Cartoon-PC-VAaudio1K<n<10K0 likes57 downloads5mo agoHugging Face11risashinoda /animalclap-dataset AnimalCLAP AnimalCLAP: Taxonomy-Aware Language-Audio Pretraining for Species Recognition and Trait InferenceICASSP 2026 AuthorsRisa Shinoda, Kaede Shiohara, Nakamasa Inoue, Hiroaki Santo, Fumio Okura Overview This dataset contains 701,020 animal sound recordings collected from: iNaturalist Xeno-Canto Splits HF Split Original Split Description train train Training data (URL only) validation test Validation data (URL only) test zero_shot… See the full description on the dataset page: https://huggingface.co/datasets/risashinoda/animalclap-dataset.audioaudio-classification1 likes49 downloads5mo agoHugging Face12EarthSpeciesProject /animalspeak-pseudovox AnimalSpeak Pseudovox Train-Unseen This dataset contains the train-unseen split of AnimalSpeak Pseudovox. Each example is a short, silence-trimmed, single-vocalization WAV clip plus compact per-clip metadata. It does not include generated conversations, captions, QA pairs, or MCQ answers. Rows: 346,907 Shards: 18 Maximum rows per shard: 20,000 Files data-20k/train-*.tar: WebDataset-style shards containing audio/<audio_name> WAV entries. metadata.parquet: one row per… See the full description on the dataset page: https://huggingface.co/datasets/EarthSpeciesProject/animalspeak-pseudovox.audioaudio-classification100K<n<1M0 likes43 downloads5mo agoHugging Face13SLLM-multi-hop /AnimalQA Dataset Card for SAKURA-AnimalQA This dataset contains the audio and the single/multi-hop questions/answers of the animal track of the SAKURA benchmark from Interspeech 2025 paper, "SAKURA: On the Multi-hop Reasoning of Large Audio-Language Models Based on Speech and Audio Information". The fields of the dataset are: file: The filename of the audio files. audio: The audio recordings. attribute_label: The attribute labels (i.e., the kinds of animal making the sounds) of the audio… See the full description on the dataset page: https://huggingface.co/datasets/SLLM-multi-hop/AnimalQA.audion<1K0 likes35 downloads1y agoHugging Face14DynamicSuperbPrivate /EnvironmentalSoundClassification_ESC50-Animals_TTSaudion<1K3 likes34 downloads2y agoHugging Face15APEX-SUPERB /animal_classificationaudion<1K0 likes26 downloads1y agoHugging Face16mteb /Human-Animal-Cartoonaudion<1K0 likes23 downloads7mo agoHugging Face17jytole /AnimalAudioA dataset to fine-tune the AudioLDM Audio Generation Model audion<1K1 likes20 downloads3y agoHugging Face18windcrossroad /AnimalQA-gemini-1.5-pro-caption Dataset Card for "AnimalQA-gemini-1.5-pro-caption" More Information needed audion<1K0 likes18 downloads2y agoHugging Face19DynamicSuperb /AnimalClassification_WaveSource-Testaudion<1K0 likes17 downloads2y agoHugging Face20windcrossroad /AnimalQA-gemini-1.5-flash-fix Dataset Card for "AnimalQA-gemini-1.5-flash-fix" More Information needed audion<1K0 likes16 downloads2y agoHugging Face21windcrossroad /AnimalQA-gemini-1.5-pro Dataset Card for "AnimalQA-gemini-1.5-pro" More Information needed audion<1K0 likes15 downloads2y agoHugging Face22windcrossroad /AnimalQA-gemini-1.5-flash-caption Dataset Card for "AnimalQA-gemini-1.5-flash-caption" More Information needed audion<1K0 likes15 downloads2y agoHugging Face23windcrossroad /AnimalQA-LTUAS-caption Dataset Card for "AnimalQA-LTUAS-caption" More Information needed audion<1K0 likes15 downloads2y agoHugging Face24windcrossroad /AnimalQA-LTUAS Dataset Card for "AnimalQA-LTUAS" More Information needed audion<1K0 likes14 downloads2y agoHugging Face25Bradarr /audio-alpaca-animalsaudio1K<n<10K0 likes14 downloads2y agoHugging Face26windcrossroad /AnimalQA-gemini-1.5-flash Dataset Card for "AnimalQA-gemini-1.5-flash" More Information needed audion<1K0 likes13 downloads2y agoHugging Face27windcrossroad /AnimalQA-GPT4o Dataset Card for "AnimalQA-GPT4o" More Information needed audion<1K0 likes10 downloads2y agoHugging Face28windcrossroad /AnimalQA-GPT4o-caption Dataset Card for "AnimalQA-GPT4o-caption" More Information needed audion<1K0 likes7 downloads2y agoHugging Face29ericholam /AnimalQA-Best-Balanced-Distractorsaudion<1K0 likes7 downloads10mo agoHugging Face30GaspardNW /Animaux_RAW_AUDIOaudion<1K0 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.