andyc03/latent-policy-guard-40k
Latent Policy Guard β Training Set (40k) This is the training distribution for Latent Policy Guard (LPG) β a guardrail model that performs semantic latent deliberation over dynamic safety policies. Each record pairs an indexed policy list and a content snippet with teacher-grounded reasoning over the user's intent and the risk of policy violation, terminating in a compact verdict anchored to violated policy indices. π Paper: LPG: Balancing Efficiency and Policy Reasoning inβ¦ See the full description on the dataset page: https://huggingface.co/datasets/andyc03/latent-policy-guard-40k.
This repository belongs to andyc03 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
