CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01deepset /prompt-injections Dataset Card for "deberta-v3-base-injection-dataset" More Information needed textn<1K183 likes11k downloads2y agoHugging Face02neuralchemy /Prompt-injection-dataset advance dataset if you want for llm security https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/Prompt-injection-dataset.texttext-classification10K<n<100K30 likes2.8k downloads5mo agoHugging Face03facebook /cyberseceval3-visual-prompt-injection Dataset Card for CyberSecEval 3 - Visual Prompt Injection Benchmark Dataset Details Dataset Description This dataset provides a multimodal benchmark for visual prompt injection, with text/image inputs. It is part of CyberSecEval 3, the third edition of Meta's flagship suite of security benchmarks for LLMs to measure cybersecurity risks and capabilities across multiple domains. Language(s): English License: MIT Dataset Sources Repository: Link… See the full description on the dataset page: https://huggingface.co/datasets/facebook/cyberseceval3-visual-prompt-injection.imagetext-generation1K<n<10K10 likes2.7k downloads2y agoHugging Face04xTRam1 /safe-guard-prompt-injectionWe formulated the prompt injection detector problem as a classification problem and trained our own language model to detect whether a given user prompt is an attack or safe. First, to train our own prompt injection detector, we required high-quality labelled data; however, existing prompt injection datasets were either too small (on the magnitude of O(100)) or didn’t cover a broad spectrum of prompt injection attacks. To this end, inspired by the GLAN paper, we created a custom synthetic… See the full description on the dataset page: https://huggingface.co/datasets/xTRam1/safe-guard-prompt-injection.text10K<n<100K34 likes2.2k downloads2y agoHugging Face05nvidia /Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 Dataset Description: Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 is an RL dataset for training and evaluating a tool-using agent's ability to resist Indirect Prompt Injection (IPI) attacks hidden inside tool-returned environment data. In each record, the agent receives a benign user request that requires calling a read tool whose output contains an adversarial instruction disguised as legitimate domain content… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1.textreinforcement-learning1K<n<10K8 likes1.4k downloads4mo agoHugging Face06Lakera /mosscap_prompt_injection mosscap_prompt_injection This is a dataset of prompt injections submitted to the game Mosscap by Lakera. This variant of the game Gandalf was created for DEF CON 31. Note that the Mosscap levels may no longer be available in the future. Note that we release every prompt that we received, regardless of whether it truly is a prompt injection or not. There are hundrends of thousands of prompts and many of them are not actual prompt injections (people ask Mosscap all kinds of things).… See the full description on the dataset page: https://huggingface.co/datasets/Lakera/mosscap_prompt_injection.text100K<n<1M21 likes1.3k downloads2y agoHugging Face07rogue-security /prompt-injections-benchmarkgated Dataset: Qualifire Benchmark Prompt Injection(Jailbreak vs. Benign) Datasets Overview This dataset contains 5,000 prompts, each labeled as either jailbreak or benign. The dataset is designed for evaluating AI models' robustness against adversarial prompts and their ability to distinguish between safe and unsafe inputs. Dataset Structure Total Samples: 5,000 Labels: jailbreak, benign Columns: text: The input text label: The classification (jailbreak or benign)… See the full description on the dataset page: https://huggingface.co/datasets/rogue-security/prompt-injections-benchmark.text1K<n<10K45 likes1k downloads6mo agoHugging Face08reshabhs /SPML_Chatbot_Prompt_Injection SPML Chatbot Prompt Injection Dataset Arxiv Paper Introducing the SPML Chatbot Prompt Injection Dataset: a robust collection of system prompts designed to create realistic chatbot interactions, coupled with a diverse array of annotated user prompts that attempt to carry out prompt injection attacks. While other datasets in this domain have centered on less practical chatbot scenarios or have limited themselves to "jailbreaking" – just one aspect of prompt injection – our dataset… See the full description on the dataset page: https://huggingface.co/datasets/reshabhs/SPML_Chatbot_Prompt_Injection.tabulartext-classification10K<n<100K31 likes962 downloads2y agoHugging Face09jayavibhav /prompt-injection-safetytext10K<n<100K13 likes914 downloads2y agoHugging Face10S-Labs /prompt-injection-dataset Prompt Injection Detection Dataset A binary classification dataset for detecting prompt injection attacks in user inputs to LLM-based applications. Dataset Description This dataset is designed to train encoder-only models (e.g., BERT, RoBERTa, DistilBERT) to classify user inputs as either benign or prompt injection attempts. Classes Label Class Description 0 BENIGN Legitimate user queries 1 INJECTION Prompt injection attempts Features… See the full description on the dataset page: https://huggingface.co/datasets/S-Labs/prompt-injection-dataset.texttext-classification10K<n<100K8 likes742 downloads8mo agoHugging Face11yanismiraoui /prompt_injections Dataset Card for Prompt Injections by Yanis Miraoui 👋 Dataset Description This dataset of prompt injections enriches Large Language Models (LLMs) by providing task-specific examples and prompts, helping improve LLMs' performance and control their behavior. Dataset Summary This dataset contains over 1000 rows of prompt injections in multiple languages. It contains examples of prompt injections using different techniques such as: prompt leaking… See the full description on the dataset page: https://huggingface.co/datasets/yanismiraoui/prompt_injections.text1K<n<10K7 likes715 downloads4mo agoHugging Face12jayavibhav /prompt-injectiontext100K<n<1M8 likes646 downloads2y agoHugging Face13Necent /llm-jailbreak-prompt-injection-datasetgated LLM Jailbreak & Prompt-Injection Dataset A unified safety dataset combining 30+ public sources for training LLM guardrails, content moderation classifiers, and response-safety filters. Schema (orthogonal multi-label, WildGuard-style) Instead of a single binary is_dangerous, every example carries four orthogonal labels matching the structure used by AI2 WildGuard, IBM Granite Guardian, and Azure Prompt Shields: Column Type Description prompt str The user/attack… See the full description on the dataset page: https://huggingface.co/datasets/Necent/llm-jailbreak-prompt-injection-dataset.tabulartext-classification1M<n<10M39 likes620 downloads5mo agoHugging Face143nesdeniz /agentic-prompt-injection-boundary-pairs Agentic Prompt-Injection Boundary Pairs Most prompt-injection datasets make the attack easy to recognize. The malicious row contains obvious override language, while the benign row discusses something unrelated. A classifier can look capable without learning the boundary that matters in production. This dataset takes a stricter approach. Each attack is paired with a legitimate request from the same workflow. The two rows share the asset, role, tool and topic. What changes is… See the full description on the dataset page: https://huggingface.co/datasets/3nesdeniz/agentic-prompt-injection-boundary-pairs.texttext-classification1K<n<10K5 likes506 downloads2mo agoHugging Face15JasperLS /prompt-injections Dataset Card for "deberta-v3-base-injection-dataset" More Information needed textn<1K21 likes443 downloads3y agoHugging Face16joshuapenman /semantic-overlays-injection Semantic Overlays — injection training corpus The training corpus for the "do-not-execute" overlay of Semantic Overlays: Mitigating Prompt Injection with Annotations Beyond Tokens and Steering Vectors (arXiv:2608.23873), released for both base models used in the paper. paper arXiv:2608.23873 code semantic-overlays trained adapters semantic-overlays-adapters interactive demo semantic-overlays.vercel.app The companion code tokenizes these files into… See the full description on the dataset page: https://huggingface.co/datasets/joshuapenman/semantic-overlays-injection.texttext-generationn<1K0 likes368 downloads25d agoHugging Face17wambosec /prompt-injections wambosec/prompt-injections A dataset of prompts for training prompt injection detection models. Dataset Description This dataset contains prompts labeled as either benign (normal user requests) or malicious (prompt injection attacks). Dataset Statistics Total prompts: 5,766 Benign prompts: 2,340 Malicious prompts: 3,426 Malicious ratio: 59.4% Dataset Structure { "prompt": str, # The prompt text "label": int, # 0 =… See the full description on the dataset page: https://huggingface.co/datasets/wambosec/prompt-injections.texttext-classification1K<n<10K3 likes348 downloads8mo agoHugging Face18BentoUniAcc /PDF_Injection_Synthetic_Dataset# PDF Prompt Injection Detector — Dataset Course: Introduction to Data Science, 2026 Problem Statement Large Language Models are increasingly used to process PDFs — summarizing contracts, extracting data, answering questions. Attackers exploit this by hiding malicious instructions inside PDFs using techniques that are invisible to human readers but fully visible to the LLM parser receiving the extracted text. This dataset was built to study, detect, and explain these prompt… See the full description on the dataset page: https://huggingface.co/datasets/BentoUniAcc/PDF_Injection_Synthetic_Dataset.document1K<n<10K0 likes319 downloads3mo agoHugging Face193nesdeniz /turkish-conversation-prompt-injection Turkish Conversation Prompt-Injection Dataset Canonical dataset release: Hugging Face hosts the dataset viewer and downloads. This GitHub repository contains the authoring sources, release files, documentation, deterministic build pipeline and validation scripts. Version 1.0.2 has the permanent DOI 10.5281/zenodo.21379389 for its Zenodo release archive. The interactive dataset explorer provides side-by-side inspection of all 150 controlled boundary pairs, complete row… See the full description on the dataset page: https://huggingface.co/datasets/3nesdeniz/turkish-conversation-prompt-injection.texttext-classificationn<1K4 likes303 downloads2mo agoHugging Face20imoxto /prompt_injection_cleaned_dataset-v2 Dataset Card for "prompt_injection_cleaned_dataset-v2" More Information needed text100K<n<1M11 likes295 downloads3y agoHugging Face21MAlmasabi /Indirect-Prompt-Injection-BIPIA-GPTgated Indirect Prompt Injection Detection Dataset (BIPIA + GPT-4o-mini) Dataset Summary This dataset contains 70,000 examples for detecting indirect prompt injection attacks in Large Language Models. It combines: 35,000 malicious samples from the BIPIA benchmark (cleaned and processed) 35,000 benign samples generated using GPT-4o-mini Indirect prompt injection attacks embed malicious instructions within external content (code, table, email, webAQ, abstract) that LLMs process… See the full description on the dataset page: https://huggingface.co/datasets/MAlmasabi/Indirect-Prompt-Injection-BIPIA-GPT.text10K<n<100K8 likes284 downloads8mo agoHugging Face22neuralchemy /prompt-injection-Threat-Matrix CATEGORIZED DATASET - easy to use https://huggingface.co/datasets/neuralchemy/prompt-injection-dataset-categorized Neuralchemy Prompt Injection Threat Matrix A professional-grade prompt injection and jailbreak detection dataset featuring 32,320 curated samples across 5 dimensions with full threat intelligence schema including technique classification, severity scoring, attack surface detection, and ambiguity flagging. Built for training production-grade LLM… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix.tabulartext-classification10K<n<100K3 likes280 downloads3mo agoHugging Face23cgoosen /prompt_injection_password_or_secrettextn<1K3 likes270 downloads3y agoHugging Face24AltaySec /turkish-llm-injection 🇹🇷 AltaySec Turkish LLM Prompt Injection Dataset (v0.2) Türkiye'nin ilk Türkçe-öncelikli, kategorize edilmiş LLM prompt injection veri seti — genişletilmiş sürüm. 📌 TL;DR 300 elle/üretim-destekli hazırlanmış Türkçe prompt injection payload'u, 12 saldırı kategorisi × 25, OWASP LLM Top 10 (2025) ile eşlenmiş. v0.1'in 120 çekirdek payload'una, AltayDuel arenasındaki bulgular ışığında üretilip düşmanca kalite/dedup denetiminden geçirilmiş 180 yeni payload eklendi.… See the full description on the dataset page: https://huggingface.co/datasets/AltaySec/turkish-llm-injection.texttext-classificationn<1K2 likes257 downloads1mo agoHugging Face25hirundo-io /prompt-injection-purple-llamatextn<1K0 likes256 downloads1y agoHugging Face26Radda /prompt-injection-dataset-v2 Prompt Injection Detection Dataset v2 Description Binary classification dataset for detecting prompt injection attacks in LLM inputs. Label 1 = injection, label 0 = benign. Composition Split Rows Injection % Description train.json 24,196 57.4% Training set (relabeled via 3-method consensus) val.json 2,688 57.4% Validation set test_clean.json 749 56.7% Leakage-free test set (near-duplicates with train removed) test_original.json 942… See the full description on the dataset page: https://huggingface.co/datasets/Radda/prompt-injection-dataset-v2.text-classification10K<n<100K2 likes252 downloads26d agoHugging Face27zachz /prompt-injection-benchmark Prompt Injection Benchmark A curated dataset of labeled prompt injection attacks and benign prompts for testing and benchmarking injection detection systems. Dataset Description This dataset contains 200 examples across 7 attack categories, plus 100 benign prompts. Each example is labeled with: text: The prompt text label: injection or benign category: Attack category (e.g., instruction_override, role_hijack) severity: low, medium, high, or critical Attack… See the full description on the dataset page: https://huggingface.co/datasets/zachz/prompt-injection-benchmark.texttext-classificationn<1K1 likes247 downloads6mo agoHugging Face28neuralchemy /prompt-injection-dataset-categorized Prompt Injection Dataset — Categorized (Threat Matrix V2) Welcome to Prompt Injection Dataset – Categorized (formerly Threat Matrix), by Neuralchemy. This is the successor to our original Prompt Injection Threat Matrix dataset. Instead of one multi-label table, this version splits the taxonomy into 7 clean, single-purpose subsets — 6 taxonomy dimensions plus a bonus ambiguity flag — so you can train a focused specialist model on each one instead of fighting multi-task learning.… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/prompt-injection-dataset-categorized.tabulartext-classification100K<n<1M1 likes241 downloads2mo agoHugging Face293nesdeniz /agentic-prompt-injection-5k Agentic Prompt-Injection 5K 5,000 examples of agentic and indirect prompt injection. A curated, paired benign/attack dataset for evaluating and training prompt-injection detectors and LLM guardrails. It focuses on the harder, agentic surface: tool/function abuse, RAG-document-embedded (indirect) injection, memory and trust-boundary poisoning, and approval/authority escalation. Curated by Enes Deniz (ORCID 0009-0006-9491-3565), Co-Founder at AltaySec and OWASP AI Exchange / GenAI… See the full description on the dataset page: https://huggingface.co/datasets/3nesdeniz/agentic-prompt-injection-5k.tabulartext-classification1K<n<10K3 likes240 downloads1mo agoHugging Face30imoxto /prompt_injection_cleaned_dataset Dataset Card for "prompt_injection_cleaned_dataset" More Information needed tabular100K<n<1M6 likes238 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.