CoolFace
Datasetpublic

andyc03/latent-policy-guard-40k

Latent Policy Guard β€” Training Set (40k) This is the training distribution for Latent Policy Guard (LPG) β€” a guardrail model that performs semantic latent deliberation over dynamic safety policies. Each record pairs an indexed policy list and a content snippet with teacher-grounded reasoning over the user's intent and the risk of policy violation, terminating in a compact verdict anchored to violated policy indices. πŸ“„ Paper: LPG: Balancing Efficiency and Policy Reasoning in… See the full description on the dataset page: https://huggingface.co/datasets/andyc03/latent-policy-guard-40k.

sourceHugging Faceotherupdated 4mo agoView on Hugging Face
3likes42downloads
4 commits on main
9bc19804mo ago

Fix source-dataset links: DynaBench (montehoover/DynaBench), GuardSet-X (AI-Secure/PolyGuard)

andyc03
0f2f4ea4mo ago

Add dataset card

andyc03
9b0e1ee4mo ago

Add LPG 40k training set (cleaned: training fields only)

andyc03
92ab4e24mo ago

initial commit

andyc03