CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01rubend18 /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K275 likes24k downloads3y agoHugging Face02JailbreakV-28K /JailBreakV-28k ⛓‍💥 JailBreakV-28K: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks 🌐 GitHub | 🛎 Project Page | 👉 Download full datasets If you like our project, please give us a star ⭐ on Hugging Face for the latest update. 📰 News Date Event 2024/07/09 🎉 Our paper is accepted by COLM 2024. 2024/06/22 🛠️ We have updated our version to V0.2, which supports users to customize their attack models… See the full description on the dataset page: https://huggingface.co/datasets/JailbreakV-28K/JailBreakV-28k.imagetext-generation10K<n<100K72 likes21k downloads2y agoHugging Face03TrustAIRLab /in-the-wild-jailbreak-prompts In-The-Wild Jailbreak Prompts on LLMs This is the official repository for the ACM CCS 2024 paper "Do Anything Now'': Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models by Xinyue Shen, Zeyuan Chen, Michael Backes, Yun Shen, and Yang Zhang. In this project, employing our new framework JailbreakHub, we conduct the first measurement study on jailbreak prompts in the wild, with 15,140 prompts collected from December 2022 to December 2023 (including 1,405… See the full description on the dataset page: https://huggingface.co/datasets/TrustAIRLab/in-the-wild-jailbreak-prompts.tabulartext-generation10K<n<100K43 likes6.8k downloads2y agoHugging Face04AiActivity /All-Prompt-Jailbreakimagetext-generationn<1K10 likes1.5k downloads1y agoHugging Face05Mike1997126 /All-Prompt-Jailbreakimagetext-generationn<1K1 likes559 downloads8mo agoHugging Face06JakkMehoffFriend /All-Prompt-Jailbreakimagetext-generationn<1K0 likes456 downloads3mo agoHugging Face07CaptainSlayAh0 /All-Prompt-Jailbreakimagetext-generationn<1K1 likes445 downloads4mo agoHugging Face08ahmedmostafa0521 /All-Prompt-Jailbreakimagetext-generationn<1K0 likes423 downloads4mo agoHugging Face09xunguangwang /JailbreakGuardrailBenchmark An Open Benchmark for Evaluating Jailbreak Guardrails in Large Language Models Introduction This repository provides instruction datasets in our SoK paper, SoK: Evaluating Jailbreak Guardrails for Large Language Models. The datasets are collected from various sources to evaluate the effectiveness of jailbreak guardrails in large language models (LLMs), including harmful prompts (i.e., JailbreakHub, JailbreakBench, MultiJail, and SafeMTData) and normal prompts (i.e.… See the full description on the dataset page: https://huggingface.co/datasets/xunguangwang/JailbreakGuardrailBenchmark.tabulartext-generation1K<n<10K5 likes232 downloads11mo agoHugging Face10OnerAYTAS /Turkish_prompt_injection_jailbreak_dataset [!NOTE] Türkçe Prompt Injection & Jailbreak Veri Seti 📌 Atıf / Citation Bu veri setini akademik çalışmalarda, model değerlendirmelerinde, güvenlik analizlerinde veya türev araştırmalarda kullanırsanız lütfen aşağıdaki makaleye atıf veriniz: Aytaş, Ö.; Şen, T.; Diri, B.; Biricik, G.; Bayram, M.A. Benchmarking Prompt Injection Attacks on LLMs: Turkish Vulnerability Assessment and English Comparative Analysis. Applied Sciences 2026, 16(13), 6740.… See the full description on the dataset page: https://huggingface.co/datasets/OnerAYTAS/Turkish_prompt_injection_jailbreak_dataset.text-generation10K<n<100K3 likes210 downloads3mo agoHugging Face11Mindgard /evaded-prompt-injection-and-jailbreak-samplesgatedThis dataset originates from our paper 'Bypassing Prompt Injection and Jailbreak Detection in LLM Guardrails'. The dataset contains a mixture of prompt injections and jailbreak samples modified via character injection and adversarial ML evasion techniques (Techniques can be found within the paper above). For each sample we provide the original unaltered prompt and a modified prompt, the attack_name outlines which attack technique was used to modify the sample. Acknowledgements… See the full description on the dataset page: https://huggingface.co/datasets/Mindgard/evaded-prompt-injection-and-jailbreak-samples.texttext-classification10K<n<100K20 likes161 downloads1y agoHugging Face12nvidia /Nemotron-RL-Jailbreak-Robustness-v1 Dataset Description: The Nemotron-RL-Jailbreak-Robustness-v1 data is designed to (1) strengthen model robustness against a variety of adversarial jailbreak techniques and (2) at the same time improve adherence to behavioral policies. This dataset is a collection of hybrid (open-source and synthetically generated) collection of adversarial prompts designed to elicit undesirable behavior from large language models. That's it, just prompts, responses are generated during training… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Jailbreak-Robustness-v1.textreinforcement-learning1K<n<10K1 likes118 downloads4mo agoHugging Face13GA-Res /AI-Jailbreak-Prompts Dataset Card for Dataset Name Name Jailbreak Prompts Dataset Summary Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [German] tabularquestion-answeringn<1K2 likes106 downloads5mo agoHugging Face14Ngixdev /JailBreakV-28k ⛓‍💥 JailBreakV-28K: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks 🌐 GitHub | 🛎 Project Page | 👉 Download full datasets If you like our project, please give us a star ⭐ on Hugging Face for the latest update. 📰 News Date Event 2024/07/09 🎉 Our paper is accepted by COLM 2024. 2024/06/22 🛠️ We have updated our version to V0.2, which supports users to customize their attack models… See the full description on the dataset page: https://huggingface.co/datasets/Ngixdev/JailBreakV-28k.imagetext-generation10K<n<100K1 likes102 downloads6mo agoHugging Face15while-ai /airline-resist-jailbreaks airline-resist-jailbreaks Made with the whileai SDK · Collection: Robustness Jailbreak resistance for a customer support agent, trained on simulated attacks and tested on real ones. The real attacks come from elder-plinius/L1B3RT4S, a public library of working jailbreaks. We read it to extract the attack techniques and never trained on a single string from it. It is the evaluation set, unseen by the model. On 165 unseen blocks from a public jailbreak library the agent holds its… See the full description on the dataset page: https://huggingface.co/datasets/while-ai/airline-resist-jailbreaks.texttext-generationn<1K0 likes98 downloads2d agoHugging Face16h4sch /AI-Jailbreak-Prompts Dataset Card for Dataset Name Name Jailbreak Prompts Dataset Summary Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [German] tabularquestion-answeringn<1K4 likes90 downloads9mo agoHugging Face17groupfairnessllm /r1-1776-jailbreak R1-1776 Jailbreaking Examples The R1-1776 Jailbreaking Examples dataset comprises instances where attempts were made to bypass the safety mechanisms of the R1-1776 model—a version of DeepSeek-R1 fine-tuned by Perplexity AI to eliminate specific censorship while maintaining robust reasoning capabilities. This dataset serves as a resource for analyzing vulnerabilities in language models and developing strategies to enhance their safety and reliability. Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/groupfairnessllm/r1-1776-jailbreak.textquestion-answeringn<1K7 likes85 downloads2y agoHugging Face18XxXNebuCHADnezzarXxx /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes76 downloads8mo agoHugging Face19daveiloper /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes74 downloads9mo agoHugging Face20philosopher-from-god /ChatGPT-Jailbreak-Prompts-rubend18 Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K2 likes73 downloads1y agoHugging Face21MarkrAI /ko-jailbreak 🔓 ko-jailbreak 공개된 영어 jailbreak(탈옥) 벤치마크들을 한국어로 번역·정제하고 소스 간 중복을 제거해 하나로 통합한 한국어 jailbreak 평가 벤치마크입니다. 한국어 LLM·안전 가드가 실제 탈옥 공격에 얼마나 견고한지 측정합니다. 💻 GitHub (코드·파이프라인·리포트): https://github.com/Marker-Inc-Korea/KO-JailBreak 📊 규모: 11,441개 (behavior 10,303 + template 1,138) 새로운 공격 기법을 제안하지 않습니다. 이미 검증된 공개 영어 jailbreak 데이터셋들을 재배포가 허용되는 라이선스(MIT/Apache-2.0)에 한해 선별하고, 한국어로 번역·품질 검수한 뒤 소스 간 중복을 제거해 한국어 환경에서 바로 쓸 수 있는 단일 벤치마크로 재구성했습니다. 데이터 구성 jailbreak 벤치마크의 두 축을 함께 담아… See the full description on the dataset page: https://huggingface.co/datasets/MarkrAI/ko-jailbreak.text-generation10K<n<100K3 likes71 downloads3mo agoHugging Face22Cefress /in-the-wild-jailbreak-prompts In-The-Wild Jailbreak Prompts on LLMs This is the official repository for the ACM CCS 2024 paper "Do Anything Now'': Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models by Xinyue Shen, Zeyuan Chen, Michael Backes, Yun Shen, and Yang Zhang. In this project, employing our new framework JailbreakHub, we conduct the first measurement study on jailbreak prompts in the wild, with 15,140 prompts collected from December 2022 to December 2023 (including 1… See the full description on the dataset page: https://huggingface.co/datasets/Cefress/in-the-wild-jailbreak-prompts.tabulartext-generation10K<n<100K0 likes71 downloads8d agoHugging Face23nikitaapawar /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K0 likes67 downloads13d agoHugging Face24daveiloper /AI-Jailbreak-Prompts Dataset Card for Dataset Name Name Jailbreak Prompts Dataset Summary Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [German] tabularquestion-answeringn<1K3 likes65 downloads9mo agoHugging Face25zorasoie /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K0 likes62 downloads15d agoHugging Face26HentaiLovers911 /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K2 likes61 downloads6mo agoHugging Face27Ngixdev /in-the-wild-jailbreak-prompts In-The-Wild Jailbreak Prompts on LLMs This is the official repository for the ACM CCS 2024 paper "Do Anything Now'': Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models by Xinyue Shen, Zeyuan Chen, Michael Backes, Yun Shen, and Yang Zhang. In this project, employing our new framework JailbreakHub, we conduct the first measurement study on jailbreak prompts in the wild, with 15,140 prompts collected from December 2022 to December 2023 (including 1,405… See the full description on the dataset page: https://huggingface.co/datasets/Ngixdev/in-the-wild-jailbreak-prompts.tabulartext-generation10K<n<100K1 likes60 downloads6mo agoHugging Face28nahsa /in-the-wild-jailbreak-prompts In-The-Wild Jailbreak Prompts on LLMs This is the official repository for the ACM CCS 2024 paper "Do Anything Now'': Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models by Xinyue Shen, Zeyuan Chen, Michael Backes, Yun Shen, and Yang Zhang. In this project, employing our new framework JailbreakHub, we conduct the first measurement study on jailbreak prompts in the wild, with 15,140 prompts collected from December 2022 to December 2023 (including 1… See the full description on the dataset page: https://huggingface.co/datasets/nahsa/in-the-wild-jailbreak-prompts.tabulartext-generation10K<n<100K0 likes60 downloads8d agoHugging Face29Tman9099 /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes59 downloads8mo agoHugging Face30RAED1ax /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes59 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.