CoolFace
Datasetpublic

3nesdeniz/agentic-prompt-injection-boundary-pairs

Agentic Prompt-Injection Boundary Pairs Most prompt-injection datasets make the attack easy to recognize. The malicious row contains obvious override language, while the benign row discusses something unrelated. A classifier can look capable without learning the boundary that matters in production. This dataset takes a stricter approach. Each attack is paired with a legitimate request from the same workflow. The two rows share the asset, role, tool and topic. What changes is… See the full description on the dataset page: https://huggingface.co/datasets/3nesdeniz/agentic-prompt-injection-boundary-pairs.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
6likes583downloads
5 commits on main
a5682e72mo ago

docs: add Zenodo DOI and pin reproducible build

3nesdeniz
8404bd73mo ago

Link interactive boundary explorer

3nesdeniz
9f47f3a3mo ago

docs: link dataset design article

3nesdeniz
a946e8e3mo ago

Release Agentic Prompt-Injection Boundary Pairs v1.0.0

3nesdeniz
73a1d853mo ago

initial commit

3nesdeniz