andyc03/latent-policy-guard-40k
Latent Policy Guard β Training Set (40k) This is the training distribution for Latent Policy Guard (LPG) β a guardrail model that performs semantic latent deliberation over dynamic safety policies. Each record pairs an indexed policy list and a content snippet with teacher-grounded reasoning over the user's intent and the risk of policy violation, terminating in a compact verdict anchored to violated policy indices. π Paper: LPG: Balancing Efficiency and Policy Reasoning inβ¦ See the full description on the dataset page: https://huggingface.co/datasets/andyc03/latent-policy-guard-40k.
Fix source-dataset links: DynaBench (montehoover/DynaBench), GuardSet-X (AI-Secure/PolyGuard)
Add dataset card
Add LPG 40k training set (cleaned: training fields only)
initial commit
