CoolFace
Datasetpublic

laion/captioned-ai-music-snippets

Dataset Overview A collection of short audio snippets (3–30 seconds) extracted from publicly shared Suno‑generated songs and captioned with Gemini Flash 2.0. Designed specifically to train and evaluate audio captioning models. Source Clips are randomly cut from the songs referenced in the nyuuzyou/suno repository. Captioning All excerpts have been annotated using Gemini Flash 2.0 for high‑quality, human‑readable audio descriptions.… See the full description on the dataset page: https://huggingface.co/datasets/laion/captioned-ai-music-snippets.

sourceHugging Faceapache-2.0updated 11mo agoView on Hugging Face
15likes2.1kdownloads
Dataset Card

Dataset Overview

A collection of short audio snippets (3–30 seconds) extracted from publicly shared Suno‑generated songs and captioned with Gemini Flash 2.0. Designed specifically to train and evaluate audio captioning models.

Source

Clips are randomly cut from the songs referenced in the nyuuzyou/suno repository.

Captioning

All excerpts have been annotated using Gemini Flash 2.0 for high‑quality, human‑readable audio descriptions.

License

Apache 2.0