AE-W/generative-sound-masking-input-noise-full-v1
Generative Sound Masking input-noise pool v1 This WebDataset contains 48,840 mono 16-kHz, 10.24-second input-noise clips across 49 tar shards. It combines the complete Yiming SONYC, TAU Urban Acoustic Scenes, and UrbanSound baseline with subject-balanced BABYCRY-UJM-AXA and NOTSOFAR-1 train windows. Stable sample metadata are in metadata/noise_index.jsonl; JSON beside each WAV adds hashes computed during packaging. The source datasets carry different licenses. In particular… See the full description on the dataset page: https://huggingface.co/datasets/AE-W/generative-sound-masking-input-noise-full-v1.
Generative Sound Masking input-noise pool v1
This WebDataset contains 48,840 mono 16-kHz, 10.24-second input-noise clips across 49 tar shards. It combines the complete Yiming SONYC, TAU Urban Acoustic Scenes, and UrbanSound baseline with subject-balanced BABYCRY-UJM-AXA and NOTSOFAR-1 train windows. Stable sample metadata are in metadata/noise_index.jsonl; JSON beside each WAV adds hashes computed during packaging.
The source datasets carry different licenses. In particular, UrbanSound8K is CC BY-NC 3.0, so downstream users must inspect each sample's license and redistribution_allowed fields and comply with its source terms.
