CoolFace
Datasetpublic

CaiZhiTech/Evaluation-Dataset-of-AI-Agent-Security-Guardrails

DKnownAI Agent Security Evaluation Dataset Data Fields Field Type Description text string The adversarial input (prompt) to be evaluated by a security guardrail action string Human-annotated label: blocked or allowed Citation @misc{li2026comparativeevaluationaiagent, title={A Comparative Evaluation of AI Agent Security Guardrails}, author={Qi Li and Jiu Li and Pingtao Wei and Jianjun Xu and Xueyi Wei and Jiwei Shi and… See the full description on the dataset page: https://huggingface.co/datasets/CaiZhiTech/Evaluation-Dataset-of-AI-Agent-Security-Guardrails.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
1likes100downloads
Dataset Card

DKnownAI Agent Security Evaluation Dataset

Data Fields

FieldTypeDescription
textstringThe adversarial input (prompt) to be evaluated by a security guardrail
actionstringHuman-annotated label: blocked or allowed

Citation

@misc{li2026comparativeevaluationaiagent,
      title={A Comparative Evaluation of AI Agent Security Guardrails}, 
      author={Qi Li and Jiu Li and Pingtao Wei and Jianjun Xu and Xueyi Wei and Jiwei Shi and Xuan Zhang and Yanhui Yang and Xiaodong Hui and Peng Xu and Lingquan Zhou},
      year={2026},
      eprint={2604.24826},
      archivePrefix={arXiv},
      primaryClass={cs.CR},
      url={https://arxiv.org/abs/2604.24826}, 
}

License

Apache 2.0