datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
shell-attack-evolution-dataset
Shell Honeypot Attack Request–Response Dataset
A standardized, MITRE ATT&CK–annotated dataset of post-login shell
attacks captured by Cowrie SSH/Telnet
honeypots across two collection periods — 2021–2022 and 2024. It pairs
attacker shell commands with real captured system responses, enabling both
longitudinal threat analysis and the training/evaluation of AI-driven honeypots.
This is the open-source release accompanying the paper “Unveiling Evolving
Threats: A Data Analysis… See the full description on the dataset page: https://huggingface.co/datasets/zyw-286/shell-attack-evolution-dataset.elite-attack
Elite Attack Dataset
A collection of 100 prompt injection test cases for evaluating LLM security defenses.
Dataset Description
This dataset contains prompt injection attacks designed to test the security of Large Language Models. Each attack scenario includes system messages with embedded secrets that the attacks attempt to extract.
Dataset Structure
100 test cases across 10 attack families
Verified effectiveness against Llama-3.2-3B-Instruct
Diverse attack… See the full description on the dataset page: https://huggingface.co/datasets/zyushg/elite-attack.geometry-of-harmfulness-in-multi-turn-attacks
Geometry of Harmfulness — Multi-Turn Attack Conversations
Raw multi-turn attack conversations accompanying the paper
The Geometry of Harmfulness in Multi-Turn Attacks. These are the conversations
from which the paper's hidden-state representations are extracted; the analysis
code lives in the companion repository.
Conversations were generated by running three multi-turn attack frameworks —
Crescendo, ActorAttack, and X-Teaming (attacker & judge: GPT-4o) —
against three… See the full description on the dataset page: https://huggingface.co/datasets/yelyzavetahusieva/geometry-of-harmfulness-in-multi-turn-attacks.attacker-zero-windows-v1
Attacker Zero Windows v1
This dataset is a prescored local-window derivative of
OpAI-Bench1/OpAI-Bench for the
attacker-zero Verifiers environment.
Each row contains one human / AI-aided / human sentence window from OpAI-Bench:
[Previous]: human sentence
[TARGET]: AI-aided sentence
[Next]: human sentence
The dataset intentionally stores raw window fields and deterministic detector
scores, not prompts. The environment owns prompt rendering, action formatting,
turn logic, and… See the full description on the dataset page: https://huggingface.co/datasets/oliveirabruno01/attacker-zero-windows-v1.rr-circuit-breakers-attack-completions
RR (Circuit Breakers) attack completions with three-judge scores
This dataset bundles attack completions generated against
GraySwanAI/Llama-3-8B-Instruct-RR
(the "circuit breakers" defense), each scored by three independent judges:
local:strongreject (Lin et al., StrongREJECT classifier — most permissive)
local:harmbench (HarmBench classifier — middle)
local:gpt_oss (gpt-oss-safeguard-20b — strictest)
Headline finding: judges DISAGREE dramatically on… See the full description on the dataset page: https://huggingface.co/datasets/samuelsimko/rr-circuit-breakers-attack-completions.
