datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
production-ai-guardrail-evals
Production AI Guardrail Evals
Eighteen synthetic, assertion-bearing cases for testing whether a language model can stay inside an advisory role. The cases cover roadmap intake, release readiness, and catalog-change review—the same workload families I use in the Winwood AI Toolkit around IEM Rig.
This is a sanitized public derivative, not a dump of application logs or the private evaluation corpus.
What each row contains
case_id: stable public identifier;… See the full description on the dataset page: https://huggingface.co/datasets/mattwinwood/production-ai-guardrail-evals.veto-guardrail-30k
Veto Guardrail 30K
A specialized dataset for training AI security guardrail models to evaluate tool calls against configurable security policies.
Model
This dataset was created to train Veto Warden 4B — a fast, specialized model for real-time tool call security validation.
Overview
Metric
Value
Examples
30,000
Format
Conversational (ShareGPT)
Task
Security policy evaluation
Domains
8 specialized categories
Task Description
The… See the full description on the dataset page: https://huggingface.co/datasets/ycaleb/veto-guardrail-30k.
