ProtectSkills/MaliciousAgentSkillsBench
MaliciousAgentSkillsBench A security benchmark dataset of Claude Code Agent Skills, from the USENIX Security 2026 paper "Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild β the first systematic study of malicious skills in the Claude Code ecosystem. π Paper: https://arxiv.org/abs/2602.06547 π» Code & full evaluation framework: https://github.com/protectskills/MaliciousAgentSkillsBench π¦ Permanent archive:β¦ See the full description on the dataset page: https://huggingface.co/datasets/ProtectSkills/MaliciousAgentSkillsBench.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone elseβs repository from here would need an authorised integration and the account holderβs consent, so the link goes to the source instead.
Open discussions on Hugging Face