CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01deepset /prompt-injections Dataset Card for "deberta-v3-base-injection-dataset" More Information needed textn<1K183 likes13k downloads2y agoHugging Face02neuralchemy /Prompt-injection-dataset advance dataset if you want for llm security https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/Prompt-injection-dataset.texttext-classification10K<n<100K30 likes2.7k downloads5mo agoHugging Face03rogue-security /prompt-injections-benchmarkgated Dataset: Qualifire Benchmark Prompt Injection(Jailbreak vs. Benign) Datasets Overview This dataset contains 5,000 prompts, each labeled as either jailbreak or benign. The dataset is designed for evaluating AI models' robustness against adversarial prompts and their ability to distinguish between safe and unsafe inputs. Dataset Structure Total Samples: 5,000 Labels: jailbreak, benign Columns: text: The input text label: The classification (jailbreak or benign)… See the full description on the dataset page: https://huggingface.co/datasets/rogue-security/prompt-injections-benchmark.text1K<n<10K45 likes1.1k downloads6mo agoHugging Face04jayavibhav /prompt-injection-safetytext10K<n<100K13 likes963 downloads2y agoHugging Face05jayavibhav /prompt-injectiontext100K<n<1M8 likes719 downloads2y agoHugging Face06JasperLS /prompt-injections Dataset Card for "deberta-v3-base-injection-dataset" More Information needed textn<1K21 likes433 downloads3y agoHugging Face07wambosec /prompt-injections wambosec/prompt-injections A dataset of prompts for training prompt injection detection models. Dataset Description This dataset contains prompts labeled as either benign (normal user requests) or malicious (prompt injection attacks). Dataset Statistics Total prompts: 5,766 Benign prompts: 2,340 Malicious prompts: 3,426 Malicious ratio: 59.4% Dataset Structure { "prompt": str, # The prompt text "label": int, # 0 =… See the full description on the dataset page: https://huggingface.co/datasets/wambosec/prompt-injections.texttext-classification1K<n<10K3 likes360 downloads8mo agoHugging Face08imoxto /prompt_injection_cleaned_dataset-v2 Dataset Card for "prompt_injection_cleaned_dataset-v2" More Information needed text100K<n<1M11 likes300 downloads3y agoHugging Face09hirundo-io /prompt-injection-purple-llamatextn<1K0 likes256 downloads1y agoHugging Face10imoxto /prompt_injection_cleaned_dataset Dataset Card for "prompt_injection_cleaned_dataset" More Information needed tabular100K<n<1M6 likes222 downloads3y agoHugging Face11geekyrakshit /prompt-injection-datasetCollected from the following datasets: deepset/prompt-injections xTRam1/safe-guard-prompt-injection jayavibhav/prompt-injection text100K<n<1M8 likes192 downloads2y agoHugging Face12ivanleomk /prompt_injection_password Dataset Card for "prompt_injection_password" More Information needed textn<1K1 likes157 downloads3y agoHugging Face13rikka-snow /prompt-injection-multilingualtexttext-classification1K<n<10K1 likes144 downloads2y agoHugging Face14cyberec /Prompt-injection-dataset Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29 attack categories including cutting-edge 2025 techniques Severity labels, source tracking, augmentation flags on every row… See the full description on the dataset page: https://huggingface.co/datasets/cyberec/Prompt-injection-dataset.texttext-classification10K<n<100K2 likes127 downloads6mo agoHugging Face15anggiatm /prompt-injection-gemmatext1K<n<10K0 likes108 downloads2y agoHugging Face16imoxto /prompt_injection_hackaprompt_gpt35 Dataset Card for "prompt_injection_hackaprompt_gpt35" More Information needed text100K<n<1M7 likes105 downloads3y agoHugging Face17yoyo221905 /prompt-injection-datasetMLTtext1K<n<10K0 likes102 downloads28d agoHugging Face18Octavio-Santana /prompt-injection-attack-detection-multilingual Prompt Injection Attack Detection Multilingual Dataset 📌 Overview This dataset is a merged and cleaned combination of two publicly available datasets for prompt injection detection: PromptInjectionDataset/Injection-Attack-Detection-Dataset rikka-snow/prompt-injection-multilingual The goal of this merged dataset is to provide a larger and more diverse benchmark for binary classification of prompt injection attacks. 🎯 Task Binary classification: 0 →… See the full description on the dataset page: https://huggingface.co/datasets/Octavio-Santana/prompt-injection-attack-detection-multilingual.texttext-classification1K<n<10K2 likes81 downloads7mo agoHugging Face19Eric12132 /prompt-injections-benchmark-chinesetext1K<n<10K0 likes72 downloads8mo agoHugging Face20Shomi28 /prompt-injection-dataset 🛡️ Prompt Injection Detection Dataset A comprehensive, balanced dataset of prompt injection attacks and safe prompts for training LLM security classifiers. Author: Soham DahivalkarLicense: MITTask: Binary Text Classification (safe vs injection) Dataset Description This dataset contains labeled prompts for training models to detect prompt injection attacks — the #1 vulnerability in LLM applications (OWASP LLM Top 10). Injection Categories Covered… See the full description on the dataset page: https://huggingface.co/datasets/Shomi28/prompt-injection-dataset.texttext-classification1K<n<10K0 likes60 downloads4mo agoHugging Face21Albertmade /prompt-injectiontabular1K<n<10K1 likes58 downloads2y agoHugging Face22cowWhySo /prompt-injection-watch-dataset Prompt Injection Watch Dataset Normalized prompt-injection watch dataset built from three Hugging Face source datasets. Read this before training on it Two findings from 30 August 2026, both measured on the files in this repository. Neither was known when the dataset was first published. A random split across the pooled corpus overstates performance by a wide margin. The three contributing datasets are separable from their text alone, and their positive rates… See the full description on the dataset page: https://huggingface.co/datasets/cowWhySo/prompt-injection-watch-dataset.texttext-classification10K<n<100K1 likes54 downloads24d agoHugging Face23hlyn-labs /prompt-injection-judge-dataset-v2 Prompt Injection Judge Dataset - Version 2 🛡️ Dataset Description The Prompt Injection Judge Dataset (V2) is a preference-learning corpus designed exclusively for training System-2 generative AI security judges. It utilizes the ORPO (Odds Ratio Preference Optimization) algorithm to teach Large Language Models to deliberate, analyze, and definitively flag prompt injections, jailbreaks, and adversarial attacks in user inputs. This dataset was entirely rebuilt to… See the full description on the dataset page: https://huggingface.co/datasets/hlyn-labs/prompt-injection-judge-dataset-v2.text1K<n<10K2 likes53 downloads4mo agoHugging Face24DAXAAI-Research /prompt_injection_v8text100K<n<1M0 likes53 downloads3mo agoHugging Face25cyberec /Prompt-injection-dataset2 Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29 attack categories including cutting-edge 2025 techniques Severity labels, source tracking, augmentation flags on every row… See the full description on the dataset page: https://huggingface.co/datasets/cyberec/Prompt-injection-dataset2.texttext-classification10K<n<100K1 likes52 downloads6mo agoHugging Face26enieva /Prompt-injection-dataset Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29 attack categories including cutting-edge 2025 techniques Severity labels, source tracking, augmentation flags on every row… See the full description on the dataset page: https://huggingface.co/datasets/enieva/Prompt-injection-dataset.texttext-classification10K<n<100K0 likes52 downloads5mo agoHugging Face27Weni /prompt-injections-1.0.0textn<1K1 likes49 downloads2y agoHugging Face28wambosec /prompt-injections-subtle wambosec/prompt-injections-subtle A dataset of prompts for training prompt injection detection models. Dataset Description This dataset contains prompts labeled as either benign (normal user requests) or malicious (prompt injection attacks). Dataset Statistics Total prompts: 933 Benign prompts: 267 Malicious prompts: 666 Malicious ratio: 71.4% Dataset Structure { "prompt": str, # The prompt text "label": int, # 0 =… See the full description on the dataset page: https://huggingface.co/datasets/wambosec/prompt-injections-subtle.texttext-classificationn<1K1 likes46 downloads8mo agoHugging Face29Sahana28 /Prompt-injection-dataset advance dataset if you want for llm security https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers… See the full description on the dataset page: https://huggingface.co/datasets/Sahana28/Prompt-injection-dataset.texttext-classification10K<n<100K0 likes45 downloads2mo agoHugging Face30MistyozAI /prompt-injection-15k prompt-injection-15k A dataset for training prompt injection / jailbreak detection classifiers. 15,000 examples labeled as either a prompt injection attempt (1) or a normal, benign message (0). Includes direct overrides, roleplay jailbreaks, system prompt extraction attempts, obfuscated attacks, multilingual examples, and benign edge cases (imperative tasks, casual use of words like "ignore" or "forget", short or minimal inputs). Fields text: the user message… See the full description on the dataset page: https://huggingface.co/datasets/MistyozAI/prompt-injection-15k.text10K<n<100K2 likes40 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.