datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
prompt-injections
Dataset Card for "deberta-v3-base-injection-dataset"
More Information needed
Prompt-injection-dataset
advance dataset if you want for llm security
https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix
Prompt Injection & Jailbreak Detection Dataset
A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models.
Zero data leakage — group-aware splitting confirmed
Balanced classes — ~60% malicious / 40% benign
Two configs — core for classical ML, full for transformers
29… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/Prompt-injection-dataset.prompt-injections-benchmark
Dataset: Qualifire Benchmark Prompt Injection(Jailbreak vs. Benign) Datasets
Overview
This dataset contains 5,000 prompts, each labeled as either jailbreak or benign. The dataset is designed for evaluating AI models' robustness against adversarial prompts and their ability to distinguish between safe and unsafe inputs.
Dataset Structure
Total Samples: 5,000
Labels: jailbreak, benign
Columns:
text: The input text
label: The classification (jailbreak or benign)… See the full description on the dataset page: https://huggingface.co/datasets/rogue-security/prompt-injections-benchmark.prompt-injection-safetyprompt-injection-dataset
Prompt Injection Detection Dataset
A binary classification dataset for detecting prompt injection attacks in user inputs to LLM-based applications.
Dataset Description
This dataset is designed to train encoder-only models (e.g., BERT, RoBERTa, DistilBERT) to classify user inputs as either benign or prompt injection attempts.
Classes
Label
Class
Description
0
BENIGN
Legitimate user queries
1
INJECTION
Prompt injection attempts
Features… See the full description on the dataset page: https://huggingface.co/datasets/S-Labs/prompt-injection-dataset.prompt_injections
Dataset Card for Prompt Injections by Yanis Miraoui 👋
Dataset Description
This dataset of prompt injections enriches Large Language Models (LLMs) by providing task-specific examples and prompts, helping improve LLMs' performance and control their behavior.
Dataset Summary
This dataset contains over 1000 rows of prompt injections in multiple languages. It contains examples of prompt injections using different techniques such as: prompt leaking… See the full description on the dataset page: https://huggingface.co/datasets/yanismiraoui/prompt_injections.prompt-injectionprompt-injections
Dataset Card for "deberta-v3-base-injection-dataset"
More Information needed
prompt-injections
wambosec/prompt-injections
A dataset of prompts for training prompt injection detection models.
Dataset Description
This dataset contains prompts labeled as either benign (normal user requests) or malicious (prompt injection attacks).
Dataset Statistics
Total prompts: 5,766
Benign prompts: 2,340
Malicious prompts: 3,426
Malicious ratio: 59.4%
Dataset Structure
{
"prompt": str, # The prompt text
"label": int, # 0 =… See the full description on the dataset page: https://huggingface.co/datasets/wambosec/prompt-injections.prompt_injection_cleaned_dataset-v2
Dataset Card for "prompt_injection_cleaned_dataset-v2"
More Information needed
prompt-injection-Threat-Matrix
CATEGORIZED DATASET - easy to use
https://huggingface.co/datasets/neuralchemy/prompt-injection-dataset-categorized
Neuralchemy Prompt Injection Threat Matrix
A professional-grade prompt injection and
jailbreak detection dataset featuring 32,320
curated samples across 5 dimensions
with full threat intelligence schema including
technique classification, severity scoring,
attack surface detection, and ambiguity flagging.
Built for training production-grade LLM… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix.prompt_injection_password_or_secretprompt-injection-purple-llamaprompt-injection-dataset-categorized
Prompt Injection Dataset — Categorized (Threat Matrix V2)
Welcome to Prompt Injection Dataset – Categorized (formerly Threat Matrix), by Neuralchemy.
This is the successor to our original Prompt Injection Threat Matrix dataset. Instead of one multi-label table, this version splits the taxonomy into 7 clean, single-purpose subsets — 6 taxonomy dimensions plus a bonus ambiguity flag — so you can train a focused specialist model on each one instead of fighting multi-task learning.… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/prompt-injection-dataset-categorized.prompt_injection_cleaned_dataset
Dataset Card for "prompt_injection_cleaned_dataset"
More Information needed
prompt-injection-benchmark
Prompt Injection Benchmark
A curated dataset of labeled prompt injection attacks and benign prompts for testing and benchmarking injection detection systems.
Dataset Description
This dataset contains 200 examples across 7 attack categories, plus 100 benign prompts. Each example is labeled with:
text: The prompt text
label: injection or benign
category: Attack category (e.g., instruction_override, role_hijack)
severity: low, medium, high, or critical
Attack… See the full description on the dataset page: https://huggingface.co/datasets/zachz/prompt-injection-benchmark.prompt-injection-repo-dataset
Prompt Injection Repository File Dataset
A labeled dataset for detecting prompt injection attacks in repository files — code, configs, READMEs, CI/CD workflows, and documentation that AI coding agents process as context.
What This Is (and Isn't)
This dataset targets a specific threat: indirect prompt injection via repository content. When AI coding agents (Claude Code, Cursor, Copilot, Gemini CLI) clone a repo, every file becomes part of the agent's context.… See the full description on the dataset page: https://huggingface.co/datasets/prodnull/prompt-injection-repo-dataset.prompt-injection-datasetCollected from the following datasets:
deepset/prompt-injections
xTRam1/safe-guard-prompt-injection
jayavibhav/prompt-injection
prompt-injection-attack-datasetprompt_injection_password
Dataset Card for "prompt_injection_password"
More Information needed
prompt-injection-multilingualprompt-injection-bit-signatures
Status: experimental. Experiment-specific slice. Primary public dataset: scbe-aethermoore-training-data.
Prompt Injection → Bit Signatures
24,254 labeled prompts from 4 public prompt-injection datasets, each mapped through the Six Sacred Tongues bijective tokenizer from the SCBE-AETHERMOORE framework into a lossless per-prompt bit signature.
Stratified 70/15/15 train/val/test split by (source, label) so every source is represented in every split with its original label… See the full description on the dataset page: https://huggingface.co/datasets/issdandavis/prompt-injection-bit-signatures.prompt_injection_ctf_dataset_2prompt-injection-multilayerPrompt-injection-dataset
Prompt Injection & Jailbreak Detection Dataset
A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models.
Zero data leakage — group-aware splitting confirmed
Balanced classes — ~60% malicious / 40% benign
Two configs — core for classical ML, full for transformers
29 attack categories including cutting-edge 2025 techniques
Severity labels, source tracking, augmentation flags on every row… See the full description on the dataset page: https://huggingface.co/datasets/cyberec/Prompt-injection-dataset.Prompt_injection_and_Sensitive_Data_exposure_detectionprompt_injectionPrompt_Injection_PIDSprompt-injection-gemmaPrompt-Injection-Test
