datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
agent-symbolic-guardrailsThis dataset contains data associated with the paper Symbolic Guardrails for Domain-Specific Agents: Stronger Safety and Security Guarantees Without Sacrificing Utility.
Code: https://github.com/hyn0027/agent-symbolic-guardrails
Subsets
literature_review
This subset contains the metadata of the systematic literature review data. Details are discussed in Section 3 in the paper.
adversarial_MedAgentBench
This subset contains the adversarial tasks we… See the full description on the dataset page: https://huggingface.co/datasets/hyn0027D/agent-symbolic-guardrails.clinical-tpib-pathway-stability-and-risk-guardrails-v0.1What this dataset tests
Given proposed next interventionsclassify stability in the response manifoldand add a guardrail that prevents known failure patterns.
Labels
stable_move
high_variance_move
risky_move
contraindicated_move
Typical failures
repeating tolerance loops
retrial after paradoxical worsening
allowing oscillation through exposure gaps
undertreating high-risk physiology
adding noise in flat nonresponse cases
Suggested prompt wrapper
System
You evaluate… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-tpib-pathway-stability-and-risk-guardrails-v0.1.
