datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cairo-security-audits
Cairo Security Audits
A source-traceable corpus of public Cairo and Starknet security-audit metadata and normalized finding annotations.
Version 0.3.0 packages every entry in the audit inventory frozen at keep-starknet-strange/starknet-skills@17a76e8. It covers 32 accessible reports from 10 auditing firms and 286 normalized finding annotations. Eleven records are checked against rendered reports and two link to exact vulnerable/fixed commits. The release does not redistribute… See the full description on the dataset page: https://huggingface.co/datasets/starknet-ai/cairo-security-audits.Evaluation-Dataset-of-AI-Agent-Security-Guardrails
DKnownAI Agent Security Evaluation Dataset
Data Fields
Field
Type
Description
text
string
The adversarial input (prompt) to be evaluated by a security guardrail
action
string
Human-annotated label: blocked or allowed
Citation
@misc{li2026comparativeevaluationaiagent,
title={A Comparative Evaluation of AI Agent Security Guardrails},
author={Qi Li and Jiu Li and Pingtao Wei and Jianjun Xu and Xueyi Wei and Jiwei Shi and Xuan… See the full description on the dataset page: https://huggingface.co/datasets/CaiZhiTech/Evaluation-Dataset-of-AI-Agent-Security-Guardrails.ai-security-glossary
AI/LLM Security Glossary — Türkçe + English
from datasets import load_dataset
ds = load_dataset("fevziegeyurtsevenler/ai-security-glossary")
Own the Turkish AI-security vocabulary — prompt injection, jailbreak, MCP, lethal trifecta and more.
Schema
column
meaning
term_en, term_tr
term
tr_definition, en_definition
definitions
example
one example
Related AltaySec resources
🕵️ uncloak scanner:… See the full description on the dataset page: https://huggingface.co/datasets/fevziegeyurtsevenler/ai-security-glossary.terraform-s3-security-mcqai-code-security-golden
AI Code Security — Golden Set
A small, hand-labeled benchmark of code snippets — vulnerable, safe, and
needs-audit — for evaluating how well a tool detects security problems in
AI-generated code ("vibe coding"). Every case is a minimal, self-contained
example with a known, by-construction ground-truth label.
Crucially, the set is built around safe twins: many vulnerable cases are
paired with a near-identical safe variant living at the same file path. This
makes the benchmark… See the full description on the dataset page: https://huggingface.co/datasets/axyr/ai-code-security-golden.security-ai-agent
AI Security Agent Meta and Traffic Dataset in AI Agent Marketplace | AI Agent Directory | AI Agent Index from DeepNLP
This dataset is collected from AI Agent Marketplace Index and Directory at http://www.deepnlp.org, which contains AI Agents's meta information such as agent's name, website, description, as well as the monthly updated Web performance metrics, including Google,Bing average search ranking positions, Github Stars, Arxiv References, etc.
The dataset is helpful for AI… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/security-ai-agent.adaption-ai-agent-security-failures
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-ai_agent_security_failures
This dataset contains prompt-completion pairs illustrating security failures and limitations of autonomous AI agents in software development and operations contexts. The samples cover scenarios such as prompt injection attacks, unauthorized data exfiltration, destructive command execution, and the agent's inability to access local environments or missing… See the full description on the dataset page: https://huggingface.co/datasets/melanieyes/adaption-ai-agent-security-failures.viettelsecurity-ai__security-llama3.2-3b-details
Dataset Card for Evaluation run of viettelsecurity-ai/security-llama3.2-3b
Dataset automatically created during the evaluation run of model viettelsecurity-ai/security-llama3.2-3b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/viettelsecurity-ai__security-llama3.2-3b-details.adaption-vn-ai-security-audit
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-vn_ai_security_audit
This dataset contains bilingual (Vietnamese, English, and code-switched) scenarios for auditing AI agent actions in corporate IT environments. Each sample presents a specific security context involving identity management, log handling, or data access, requiring classification as either 'benign' or 'suspicious'. The entries include detailed decision rationales… See the full description on the dataset page: https://huggingface.co/datasets/melanieyes/adaption-vn-ai-security-audit.ai-security-research-logs
