watermarking
watermarking-clips
Neural Watermarking Clips
Summary
This dataset contains 31 284 short audio clips collected as an unlabeled corpus for neural audio watermarking experiments. The clips cover environmental sounds, bird vocalizations, polyphonic music with predominant instruments and synthetic but realistic jazz drums and ragtime style piano.
Source folders and file counts
ARCA23K.audio 13 470 clips
ff101bird 7 690 clips
IRMAS training 6 706 clips
WaivOps ragtime piano 1 743 clips
WaivOps… See the full description on the dataset page: https://huggingface.co/datasets/benmainbird/watermarking-clips.repro-how-good-is-post-hoc-watermarking-with-language-model-rephrasing-traces
Agent traces
Agent sessions published from a Trackio Logbook.
llm-watermarking-papers
LLM Watermarking & Copyright Detection Papers — FineSet
A research-paper dataset on LLM Watermarking & Copyright Detection Papers, assembled, deduplicated, and quality-scored by
FineSet from arXiv and Semantic Scholar.
📸 This is a dated snapshot — generated 2026-06-19.
It is not auto-updated. Research on LLM Watermarking & Copyright Detection Papers moves fast — new papers land on arXiv every
week. Want this same dataset refreshed daily, on a topic you choose? See the bottom.… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/llm-watermarking-papers.watermarking_digress
