datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ChatGPT-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
prompts.chat
a.k.a. Awesome ChatGPT Prompts
This is a Dataset Repository mirror of prompts.chat — a social platform for AI prompts.
📢 Notice
This Hugging Face dataset is a mirror. For the latest prompts, features, and community contributions, please visit:
🌐 Website: prompts.chat
📦 GitHub: github.com/f/awesome-chatgpt-prompts
About
prompts.chat is an open-source platform where users can share, discover, and collect AI prompts from the community. The project can… See the full description on the dataset page: https://huggingface.co/datasets/fka/prompts.chat.hle_prompts_07_02_25parti-prompts
Dataset Card for PartiPrompts (P2)
Dataset Summary
PartiPrompts (P2) is a rich set of over 1600 prompts in English that we release
as part of this work. P2 can be used to measure model capabilities across
various categories and challenge aspects.
P2 prompts can be simple, allowing us to gauge the progress from scaling. They
can also be complex, such as the following 67-word description we created for
Vincent van Gogh’s The Starry Night (1889):
Oil-on-canvas painting of a… See the full description on the dataset page: https://huggingface.co/datasets/nateraw/parti-prompts.text-to-image-prompts
The dataset of the most popular text-to-image prompts.
Dataset Details
Dataset Description
Curated by: kazimir.ai
Funded by [optional]: [More Information Needed]
Shared by [optional]: https://kazimir.ai
License: apache-2.0
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]
Uses
Free to use.
Dataset Structure
CSV file… See the full description on the dataset page: https://huggingface.co/datasets/Kazimir-ai/text-to-image-prompts.CCP-sensitive-prompts
CCP Sensitive Prompts
These prompts cover sensitive topics in China, and are likely to be censored by Chinese models.
ChatGPT-prompts
ChatGPT-Prompts Dataset
Description
This dataset aims to provide an evaluation data for the Language Models to come. It has been generated using LearnGPT website.
videophy2_upsampled_promptsmarketing_user_prompts_unfiltered10k_rows_cleaned_prompts
10K Rows Cleaned Prompts Dataset
Created by Aipresso LIMITED, London, UK
⚠️ IMPORTANT: By using this dataset, you agree to our Terms of Use
You must provide attribution when using this data in publications, research, or commercial products.
Dataset Overview
A chunked collection of 2.7 million cleaned English prompts, organized into 200 files of 10,000 rows each for easy processing and distributed training of language models.
📊 Dataset Statistics
Metric… See the full description on the dataset page: https://huggingface.co/datasets/Aipresso/10k_rows_cleaned_prompts.synthetic_multilingual_llm_prompts
Image generated by DALL-E. See prompt for more details
📝🌐 Synthetic Multilingual LLM Prompts
Welcome to the "Synthetic Multilingual LLM Prompts" dataset! This comprehensive collection features 1,250 synthetic LLM prompts generated using Gretel Navigator, available in seven different languages. To ensure accuracy and diversity in prompts, and translation quality and consistency across the different languages, we employed Gretel Navigator both as a generation tool and as an… See the full description on the dataset page: https://huggingface.co/datasets/gretelai/synthetic_multilingual_llm_prompts.malicious-promptsAI-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
Jailbreak Prompts
Dataset Summary
Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[German]
prompts_under_512_tokens
Under 512 Tokens Prompts Dataset
Created by Aipresso LIMITED, London, UK
⚠️ IMPORTANT: By using this dataset, you agree to our Terms of Use
Dataset Overview
Specialized collection of short-form English prompts (under 512 tokens), perfect for training models with context length constraints or faster iteration cycles.
📊 Dataset Statistics
Metric
Value
Total Files
200
Rows Per File
10,000
Total Rows
2,000,000
Token Range
1 to 511 tokens… See the full description on the dataset page: https://huggingface.co/datasets/Aipresso/prompts_under_512_tokens.malicious-prompts-minilm-embeddingsspider-sql-promptsai-writing-prompts-2025** 100 AI Writing Prompts & Marketing Hooks 2025 — A ready-to-use human readable CSV file (no coding needed option) for blogging, social media, and AI text generation projects. **
A 100-row structured dataset of AI-ready writing prompts and marketing hooks designed for blogging, copywriting, and AI text generation. Includes tone, use case, and audience metadata for fine-tuning and automation
✨ 100 AI Writing Prompts & Marketing Hooks 2025 (AI-Ready CSV for Text Generation)
A… See the full description on the dataset page: https://huggingface.co/datasets/alexandrabozarthmanagement/ai-writing-prompts-2025.chatgpt-prompts-SwedishMMA-Diffusion-NSFW-adv-prompts-benchmark
MMA-Diffusion Adversarial Prompts (Text modal attack)
The MMA-Diffusion adversarial prompts benchmark comprises 1,000 successful adversarial prompts generated by the adversarial attack methodology presented in the paper
from CVPR 2024 titled MMA-Diffusion: MultiModal Attack on Diffusion Models. This resource is intended to assist in developing and
evaluating defense mechanisms against such attacks. The adversarial prompts are capable of bypassing the image safety checker in… See the full description on the dataset page: https://huggingface.co/datasets/YijunYang280/MMA-Diffusion-NSFW-adv-prompts-benchmark.AI-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
Jailbreak Prompts
Dataset Summary
Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[German]
DREAM-Red-Teaming-Prompts
DREAM: Scalable Red Teaming for Text-to-Image Generative Systems
This dataset contains red-teaming prompts generated by the DREAM framework, as presented in the IEEE S&P 2026 paper: "DREAM: Scalable Red Teaming for Text-to-Image Generative Systems via Distribution Modeling". These prompts are designed to evaluate and stress-test the safety mechanisms of Text-to-Image (T2I) generative systems.
⚠️ Disclaimer / WarningContent Warning: This dataset contains prompts that may be… See the full description on the dataset page: https://huggingface.co/datasets/DREAM-17k/DREAM-Red-Teaming-Prompts.chatgpt-prompts-Frenchchatgpt-prompts-Estonianharmful-prompts-pt
Harmful Prompts PT-BR
harmful-prompts-pt is a Brazilian Portuguese adaptation of the
WildJailbreak dataset,
constructed to support research on the robustness of language models against
harmful and adversarial prompts in Portuguese.
This dataset was used to train and evaluate SecBERT, a Portuguese harmful
prompt classifier presented at the International Joint Conference on Neural
Networks (IJCNN). The full paper and source code are available at
[Paper] [Code].
Caution: This dataset… See the full description on the dataset page: https://huggingface.co/datasets/Edu-p/harmful-prompts-pt.DALL-E-Prompts-OpenAI-ChatGPT
Dataset Card for Dataset Name
Dataset Summary
This dataset has been generated using Prompt Generator for OpenAI's DALL-E.
Languages
English
Dataset Structure
1.000.000 Prompts
ChatGPT-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
ChatGPT-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
Prompts_for_Voice_cloning_and_TTSChatGPT-Jailbreak-Prompts-rubend18
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
ChatGPT-Jailbreak-Prompts
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
