datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Parameter-Golf-V6-Privacy-Web-Filtering
Solutions Training V6 — Privacy Filtering, Unauthorized Access Triage, and Fast Web Signal Extraction
Overview
V6 extends the V5 auxiliary-training idea into a new direction:
the model should learn to jump over noise and sensitive junk instead of reading or repeating everything.
The dataset trains a signal-first behavior for pages, emails, logs, and incident notes:
skip ads, cookie banners, footers, newsletters, and unrelated chrome,
ignore personal-data-heavy… See the full description on the dataset page: https://huggingface.co/datasets/8Planetterraforming/Parameter-Golf-V6-Privacy-Web-Filtering.Contextualized_Privacy_Defense_Trajectory
Contextualized Privacy Defense
Paper: Contextualized Privacy Defense for LLM Agents
Code: https://github.com/SALT-NLP/contextual_privacy_defense
Abstract:
Abstract LLM agents increasingly act on users’ personal information, yet existing privacy defenses remain limited in both design and adaptability. Most prior approaches rely on static or passive defenses, such as prompting and guarding. These paradigms are insufficient for supporting contextual, proactive privacy… See the full description on the dataset page: https://huggingface.co/datasets/SALT-NLP/Contextualized_Privacy_Defense_Trajectory.
