generative
generative-sound-masking-generated-energy-v7
Generative Sound Masking — fixed-background audio masking
Incrementally generated unfiltered candidates. This is not a final selected dataset.
Each background has 15 separately generated prompt–seed outputs using gain-compensated reconstruction residuals. run_config.json pins models, source pools, parameters and implementation hashes.
For multiple workers read workers/worker-NN/progress.json; each worker reports only its assigned IDs.
Global completion requires all worker… See the full description on the dataset page: https://huggingface.co/datasets/AE-W/generative-sound-masking-generated-energy-v7.generative-sound-masking-input-noise-full-v1
Generative Sound Masking input-noise pool v1
This WebDataset contains 48,840 mono 16-kHz, 10.24-second input-noise clips
across 49 tar shards. It combines the complete Yiming SONYC, TAU Urban
Acoustic Scenes, and UrbanSound baseline with subject-balanced BABYCRY-UJM-AXA
and NOTSOFAR-1 train windows. Stable sample metadata are in
metadata/noise_index.jsonl; JSON beside each WAV adds hashes computed during
packaging.
The source datasets carry different licenses. In particular… See the full description on the dataset page: https://huggingface.co/datasets/AE-W/generative-sound-masking-input-noise-full-v1.generative-sound-masking-generated-energy-v3-smokechamp_trainning_sample
Dataset samples for Champ trainning
This dataset samples is used for Champ.
Before trainning, you need to process the datasets by SMPL & DWPOSE methods. Refer to https://github.com/fudan-generative-vision/champ/blob/master/docs/data_process.md
genvsr-video-benchmarksgen-games-v9-video-pilot1
