CoolFace
Datasetpublic

andyc03/latent-policy-guard-40k

Latent Policy Guard — Training Set (40k) This is the training distribution for Latent Policy Guard (LPG) — a guardrail model that performs semantic latent deliberation over dynamic safety policies. Each record pairs an indexed policy list and a content snippet with teacher-grounded reasoning over the user's intent and the risk of policy violation, terminating in a compact verdict anchored to violated policy indices. 📄 Paper: LPG: Balancing Efficiency and Policy Reasoning in… See the full description on the dataset page: https://huggingface.co/datasets/andyc03/latent-policy-guard-40k.

sourceHugging Faceotherupdated 4mo agoView on Hugging Face
3likes42downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
andyc03/latent-policy-guard-40k · CoolFace