soundscapes
soundscapesmarrs-global-coral-reef-soundscapesThis version of the dataset has been parsed into Parquet format for easy access and experimentation.
Raw: 1 min Audio clips collected from coral reef restoration sites around the world.
Detections: 5 second clips of detected sonotypes extracted via human-in-the-loop Agile modelling as described in: Williams et al. (2025) Evidence of ecosystem process recovery across a large-scale coral reef restoration programme using AI accelerated soundscape analysis… See the full description on the dataset page: https://huggingface.co/datasets/bens-bots/marrs-global-coral-reef-soundscapes.soundscapes
SoundScapes (Embedded)
Original Source
📌 Introduction
This dataset collects the audios and annotations from the original SoundScapes dataset (laion/soundscapes).
This dataset embeds the audio and removes it using the tar files. Only the first six batches have been employed for the making of this dataset
🙏 Acknowledgement
All credits to the Laion team and the original Soundscapes teams.
advanced-soundscapes-stage-1
Advanced Soundscapes Stage 1 — Raw Components (5M)
This dataset contains Stage 1 output from the LAION Universal Audio Annotation Pipeline (UAAP) data generation plan.
Contents
5,000 shards containing 5,000,000 soundscape recipes with raw audio components
Each soundscape row includes:
recipe.json — full recipe with timeline, events, loudness, speaker IDs, overlap/density settings
spkN.flac / spkN.json — raw speech components + full source metadata
musicN.flac /… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/advanced-soundscapes-stage-1.nps-liminal-soundscapes-v0-2
nps-liminal-soundscapes-v0-2
Fork + repair of Sonic-Forage/nps-liminal-soundscapes-v0-1 (org-create permissions on our
token are restricted, so the improved version lives under TheMindExpansionNetwork; upstream keeps
the v0-1 original untouched).
Sonic-Forage Stable Audio soundscape training dataset built from verified National Park Service
public-domain sound recordings (Rocky Mountain National Park).
v0-2 is a re-release of v0-1 with layout + caption fixes (see Changelog… See the full description on the dataset page: https://huggingface.co/datasets/TheMindExpansionNetwork/nps-liminal-soundscapes-v0-2.in-the-wild-soundscapes-gemini2.5-pro
In-the-Wild Soundscapes — Gemini 2.5 Pro Annotated
In-the-wild audio segments (sourced from public YouTube videos) with temporal sound-event
segmentation and rich natural-language captions generated by Google Gemini 2.5 Pro. Each
segment is split into time-stamped events, and every event has a detailed caption describing
the content and its acoustic/technical qualities ("Universal Technical Audio Analysis").
~67,000 segments across 338 WebDataset shards (audio_batch_*.tar, ~200… See the full description on the dataset page: https://huggingface.co/datasets/laion/in-the-wild-soundscapes-gemini2.5-pro.
