CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /LAION-Audio-300Maudio100M<n<1B74 likes18k downloads2y agoHugging Face02laion /soundscapesaudio10M<n<100M7 likes15k downloads1y agoHugging Face03laion /laions_got_talent LAION's Got Talent: Generated Voice Acting Dataset Overview "LAION's Got Talent" is a generated dataset comprising voice acting samples that exhibit a wide range of emotions, vocal bursts, topics, and content. This dataset is a component of the BUD-E project, spearheaded by LAION with support from Intel. Dataset Composition The dataset includes: Emotional Diversity: Samples portraying various emotions to facilitate research in emotional recognition and… See the full description on the dataset page: https://huggingface.co/datasets/laion/laions_got_talent.audio100K<n<1M41 likes9.7k downloads2y agoHugging Face04laion /conceptual-captions-12m-webdatasetimage10K<n<100K34 likes6.5k downloads5y agoHugging Face05laion /laion-audio-previewaudio1M<n<10M11 likes4.5k downloads2y agoHugging Face06Lakonik /laion-3m WebDataset shards Each sample contains only: {image_hash}.png {image_hash}.json image1M<n<10M1 likes3.6k downloads9mo agoHugging Face07laion /laions_got_talent_rawaudio10K<n<100K7 likes3.3k downloads2y agoHugging Face08laion /Emolia Dataset Card for Emolia Dataset Description This dataset is an enhanced version of the Emilia dataset, enriched with detailed emotion annotations. The annotations were generated using models from the EmoNet suite to provide deeper insight into the emotional content of speech. This work is based on the research and models described in the blog post "Do They See What We See?". The annotations include 54 scores for each sample, covering a wide range of emotional and… See the full description on the dataset page: https://huggingface.co/datasets/laion/Emolia.audio10M<n<100M15 likes2.1k downloads10mo agoHugging Face09laion /captioned-ai-music-snippets Dataset Overview A collection of short audio snippets (3–30 seconds) extracted from publicly shared Suno‑generated songs and captioned with Gemini Flash 2.0. Designed specifically to train and evaluate audio captioning models. Source Clips are randomly cut from the songs referenced in the nyuuzyou/suno repository. Captioning All excerpts have been annotated using Gemini Flash 2.0 for high‑quality, human‑readable audio descriptions. License Apache 2.0 audio1M<n<10M15 likes1.9k downloads11mo agoHugging Face10laion /synthetic_vocal_burstsThis repository contains the vocal bursts like giggling, laughter, shouting, crying, etc. from the following repository. https://huggingface.co/datasets/sleeping-ai/Vocal-burst We captioned them using Gemini Flash Audio 2.0. This dataset contains, this dataset contains ~ 365,000 vocal bursts from all kinds of categories. It might be helpful for pre-training audio text foundation models to generate and understand all kinds of nuances in vocal bursts. audio100K<n<1M6 likes1.6k downloads2y agoHugging Face11laion /majestrino-dataaudio1M<n<10M1 likes1.5k downloads7mo agoHugging Face12laion /CS-Arxiv-PDFs-08-25text10M<n<100M2 likes1.5k downloads1y agoHugging Face13laion /laions_got_talent_enhanced_no_metadataaudio10K<n<100K0 likes978 downloads2y agoHugging Face14mkrausio /laions_got_talent_embs_only laions_got_talent Whisper Embeddings (Embeddings + Metadata Only) This dataset contains Whisper embeddings (NPY) and metadata (JSON). The original audio files (MP3) are NOT included. Embeddings computed with: mkrausio/EmoWhisper-AnS-Small-v0.1 Includes original audio: No Includes metadata: Yes (JSON) Includes embeddings: Yes (NPY) Creation date: 2025-05-11 text1M<n<10M0 likes922 downloads1y agoHugging Face15laion /Emilia-with-Emotion-Annotations4audio10M<n<100M1 likes894 downloads1y agoHugging Face16laion /laions_got_talent_german_bicodecaudio100K<n<1M0 likes868 downloads2y agoHugging Face17laion /Emilia-with-Emotion-Annotations5audio10M<n<100M3 likes738 downloads1y agoHugging Face18laion /unsupervised_peoples_speech_raw_voice_activity_detection_snippets_part_1audio100M<n<1B4 likes671 downloads1y agoHugging Face19laion /clevr-webdatasetimage1M<n<10M7 likes646 downloads4y agoHugging Face20laion /voiceclap-data VoiceCLAP Data The audio + dense-caption mixture used to train laion/voiceclap-small and laion/voiceclap-large. Each tar shard is a WebDataset of paired <key>.flac (48 kHz mono audio) + <key>.json (caption + metadata) samples. Captions and structured attribute annotations are produced automatically by a pipeline of audio-aware LLMs — Qwen-Audio, Gemini Flash 2.5, and a thinking-mode reasoning model that scores emotion under the EmoNet taxonomy plus per-clip vocal-burst, timbre… See the full description on the dataset page: https://huggingface.co/datasets/laion/voiceclap-data.audioaudio-classification1M<n<10M0 likes562 downloads5mo agoHugging Face21laion /common-voice-subset-for-clapaudion<1K1 likes521 downloads9mo agoHugging Face22laion /Emilia-with-Emotion-Annotations3audio10M<n<100M1 likes382 downloads1y agoHugging Face23laion /laions_got_talent_previewaudio1K<n<10K1 likes380 downloads2y agoHugging Face24KBlueLeaf /laion-coco-13m-taraudio10M<n<100M1 likes342 downloads1y agoHugging Face25laion /Emilia-with-Emotion-Annotations2audio10M<n<100M1 likes316 downloads1y agoHugging Face26laion /talent_plus_rl_groups_of_50_with_audiobox_scoresaudio1M<n<10M0 likes281 downloads10mo agoHugging Face27laion /majestrino-1.00-16xk5-sae-features Majestrino 1.00 SAE — Feature Audio Samples (16x, k=5) Top-2000 activating audio samples for each feature in the Majestrino 1.00 SAE. Overview Metric Value SAE Architecture 16x expansion, k=5, d_model=768 Total Features 12,288 Alive Features 10,684 Audio per Feature Up to 2,000 highest-activating Audio Format Opus (24 kbps OGG container) Total TAR Files 1069 Source Dataset laion/majestrino-data File Structure Each TAR file… See the full description on the dataset page: https://huggingface.co/datasets/laion/majestrino-1.00-16xk5-sae-features.audioaudio-classification10M<n<100M0 likes272 downloads6mo agoHugging Face28nyu-dice-lab /imagenetpp-laion-t2iDataset Card for ImageNet++'s LAION Text-to-Image Split image100K<n<1M0 likes231 downloads2y agoHugging Face29laion /emonet-face-big EmoNet-Face: A Fine-Grained, Expert-Annotated Benchmark for Facial Emotion Recognition Dataset Summary EmoNet-Face is a comprehensive benchmark suite designed to address critical gaps in facial emotion recognition (FER). Current benchmarks often have a narrow emotional spectrum, lack demographic diversity, and use uncontrolled imagery. EmoNet-Face provides a robust foundation for developing and evaluating AI systems with a deeper, more nuanced understanding of human… See the full description on the dataset page: https://huggingface.co/datasets/laion/emonet-face-big.image100K<n<1M10 likes229 downloads11mo agoHugging Face30laion /timbre-audio-caption-pairsaudio100K<n<1M2 likes206 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.