CoolFace
Datasetpublic

ProtectSkills/MaliciousAgentSkillsBench

MaliciousAgentSkillsBench A security benchmark dataset of Claude Code Agent Skills, from the USENIX Security 2026 paper "Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild β€” the first systematic study of malicious skills in the Claude Code ecosystem. πŸ“„ Paper: https://arxiv.org/abs/2602.06547 πŸ’» Code & full evaluation framework: https://github.com/protectskills/MaliciousAgentSkillsBench πŸ“¦ Permanent archive:… See the full description on the dataset page: https://huggingface.co/datasets/ProtectSkills/MaliciousAgentSkillsBench.

sourceHugging Facemitupdated 3mo agoView on Hugging Face
8likes208downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ProtectSkills/MaliciousAgentSkillsBench Β· CoolFace