datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
attackdex-paldeaSingle pokemon datasets containing all the attacks (from levelling or TMs) learnable by the relative monster. All the data refer to the Paldea region and they come from the project discussed in https://medium.com/@virtualmartire/i-built-an-algorithm-that-finds-the-optimal-pokemon-team-01ea152824a9.
Cybersecurity_Attackprompt-injection-attack-datasetNADW-network-attacks-dataset
Network Traffic Dataset for Anomaly Detection
Overview
This project presents a comprehensive network traffic dataset used for training AI models for anomaly detection in cybersecurity. The dataset was collected using Wireshark and includes both normal network traffic and various types of simulated network attacks. These attacks cover a wide range of common cybersecurity threats, providing an ideal resource for training systems to detect and respond to real-time network… See the full description on the dataset page: https://huggingface.co/datasets/onurkya7/NADW-network-attacks-dataset.multi-turn_jailbreak_attack_datasets
Multi-Turn Jailbreak Attack Datasets
Description
This dataset was created to compare single-turn and multi-turn jailbreak attacks on large language models (LLMs). The primary goal is to take a single harmful prompt and distribute the harm over multiple turns, making each prompt appear harmless in isolation. This approach is compared against traditional single-turn attacks with the complete prompt to understand their relative impacts and failure modes. The key feature of… See the full description on the dataset page: https://huggingface.co/datasets/carl213/multi-turn_jailbreak_attack_datasets.web-attacks-longcyber_MITRE_attack_tactics-and-techniquesThe dataset is question answering for MITRE tactics and techniques for version 15. Data sources are:
Tactics
Techniques
web-attacksredteaming-attack-target
Annotated version of DEFCON 31 Generative AI Red Teaming dataset with additional labels for attack targets.
This dataset is an extended version of the DEFCON31 Generative AI Red Teaming dataset, released by Humane Intelligence.
Our team conducted additional labeling on the accepted attack samples to annotate:
Attack Targets (e.g., gender, race, age, political orientation)
Attack Types (e.g., question, request, build-up, scenario assumption, misinformation injection) →… See the full description on the dataset page: https://huggingface.co/datasets/TTA01/redteaming-attack-target.attack_ICSelite-attack
Elite Attack Dataset
A collection of 100 prompt injection test cases for evaluating LLM security defenses.
Dataset Description
This dataset contains prompt injection attacks designed to test the security of Large Language Models. Each attack scenario includes system messages with embedded secrets that the attacks attempt to extract.
Dataset Structure
100 test cases across 10 attack families
Verified effectiveness against Llama-3.2-3B-Instruct
Diverse attack… See the full description on the dataset page: https://huggingface.co/datasets/zyushg/elite-attack.web-attack-detectionThe dataset contains 625,904 attack payload samples, with 294,771 labeled as 1 and 331,129 labeled as 0, including SQL injection, XSS, command injection, and other vulnerabilities.
mitre-attack-ttp-labeled-instructions
MITRE ATT&CK TTP Mapping Dataset
Training and evaluation data for mapping adversarial behavior descriptions (CTI reports,
CTF writeups, CISA advisories) to MITRE ATT&CK Tactics, Techniques, and Procedures (TTPs).
Built as my individual contribution to a research project conducted at LORIA (supervised by Jean-Yves Marion). This dataset was developed and used to fine-tune skyylord/qwen3-emb-0.6b-ttp with CachedMultipleNegativesRankingLoss and ANCE-style hard negative re-mining.… See the full description on the dataset page: https://huggingface.co/datasets/skyylord/mitre-attack-ttp-labeled-instructions.cybersecurity-attack-datasetllm-adaptive-attacksredteaming-attack-type
Annotated version of DEFCON 31 Generative AI Red Teaming dataset with additional labels for attack types.
This dataset is an extended version of the DEFCON31 Generative AI Red Teaming dataset, released by Humane Intelligence.
Our team conducted additional labeling on the accepted attack samples to annotate:
Attack Targets (e.g., gender, race, age, political orientation) → tta01/redteaming-attack-target
Attack Types (e.g., question, request, build-up, scenario assumption… See the full description on the dataset page: https://huggingface.co/datasets/TTA01/redteaming-attack-type.Cybersecurity_Attack_DatasetMed-MLLM-Attackweb-attacks-ab2NF-ToN-IoT-Attack_detailsCybersecurity_Attack_Datasetweb-attacks-abattackmcq_attackweb-attacks-oldweb-attack-detectionThe dataset contains 625,904 attack payload samples, with 294,771 labeled as 1 and 331,129 labeled as 0, including SQL injection, XSS, command injection, and other vulnerabilities.
AttackEvalLLM-Wallet-drain-attack-patternsmcq_attack_ratingcheckgpt_attack_train
