CoolFace
Datasetpublic

ProtectSkills/MaliciousAgentSkillsBench

MaliciousAgentSkillsBench A security benchmark dataset of Claude Code Agent Skills, from the USENIX Security 2026 paper "Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild โ€” the first systematic study of malicious skills in the Claude Code ecosystem. ๐Ÿ“„ Paper: https://arxiv.org/abs/2602.06547 ๐Ÿ’ป Code & full evaluation framework: https://github.com/protectskills/MaliciousAgentSkillsBench ๐Ÿ“ฆ Permanent archive:โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/ProtectSkills/MaliciousAgentSkillsBench.

sourceHugging Facemitupdated 3mo agoView on Hugging Face
8likes208downloads
4 commits on main
422bf343mo ago

Update dataset to paper final release: 98,380-skill snapshot; 157 confirmed malicious (add Severity); refresh dataset card

Yiiim0
235221c8mo ago

Update README.md

Yiiim0
c6c4c858mo ago

Upload 3 files

Yiiim0
3181e828mo ago

initial commit

Yiiim0