datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multi-turn_jailbreak_attack_datasets
Multi-Turn Jailbreak Attack Datasets
Description
This dataset was created to compare single-turn and multi-turn jailbreak attacks on large language models (LLMs). The primary goal is to take a single harmful prompt and distribute the harm over multiple turns, making each prompt appear harmless in isolation. This approach is compared against traditional single-turn attacks with the complete prompt to understand their relative impacts and failure modes. The key feature of… See the full description on the dataset page: https://huggingface.co/datasets/carl213/multi-turn_jailbreak_attack_datasets.Nemotron-RL-Instruction-Following-MultiTurnChat-v1-prompt-only
Nemotron-RL-Instruction-Following-MultiTurnChat-v1-prompt-only
Prompt-only extraction from nvidia/Nemotron-RL-Instruction-Following-MultiTurnChat-v1.
Files:
prompts.csv: one prompt extraction record per source row. Records include
prompt, separated system_prompt, and structured tools when the source row
defines available tools. Nested values are JSON-encoded inside CSV cells.
summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Instruction-Following-MultiTurnChat-v1-prompt-only.travel-multi-turn-chat-geminiillicit-general-multi-turn
Illicit General Multi-Turn Conversations
Multi-turn adversarial conversations that successfully elicited harmful illicit content from AI models. This sample dataset contains 5 conversations (52 turns) covering chemical weapons, cyber threats, and other safety-critical domains.
Dataset Statistics
Metric
Value
Conversations
5
Total Turns
52
Avg Turns/Conv
10.4
Harm Categories
3
Harm Categories
Category
Turns
Description
Chemical… See the full description on the dataset page: https://huggingface.co/datasets/GoJulyAI/illicit-general-multi-turn.ChatGPT-Style_1M_Conversation_Dataset_Multi-Turnillicit-bio-multi-turn
Illicit Bio Multi-Turn Conversations
Multi-turn adversarial conversations that successfully elicited harmful bio-safety content from AI models. This sample dataset contains 5 conversations (57 turns) covering bioweapons and related threats.
Dataset Statistics
Metric
Value
Conversations
5
Total Turns
57
Avg Turns/Conv
11.4
Harm Categories
3
Harm Categories
Category
Turns
Description
Bioweapons
34
Information about biological… See the full description on the dataset page: https://huggingface.co/datasets/GoJulyAI/illicit-bio-multi-turn.trust-and-safety-multiturn-evaluation-dataset
Dataset Description
This dataset is designed to evaluate the performance of LLM-based graders on safety-related conversations.
The dataset consists of model responses generated during multi-turn safety evaluations along with reference labels indicating whether the responses comply with safety policies.
This benchmark evaluates a grader's ability to accurately identify safe and unsafe model behavior across different safety categories.
Dataset Creation
The benchmark… See the full description on the dataset page: https://huggingface.co/datasets/CentificAIResearch/trust-and-safety-multiturn-evaluation-dataset.multi_turn-NIPS2026
Dataset Card for Scientific RAG Benchmark (Scenario)
Dataset Details
Dataset Description
This dataset is part of the Scientific RAG Benchmark Collection-NIPS2026.It is designed for evaluating Retrieval-Augmented Generation (RAG) systems and large language models on domain-specific scientific question-answering tasks.
Each scenario contains expert-curated question–answer pairs grounded in peer-reviewed scientific literature, with explicit DOI references to… See the full description on the dataset page: https://huggingface.co/datasets/anonymousauthor2026nips/multi_turn-NIPS2026.MultiTurnDialogueSets
MultiTurnDialogueSets
tags: multi-turn dialogue, turn-taking, machine learning
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'MultiTurnDialogueSets' dataset is a curated collection of conversation exchanges designed for training and evaluating machine learning models on multi-turn dialogue systems. Each entry in the dataset captures the flow of a conversation, which includes a set of utterances from a user and corresponding… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/MultiTurnDialogueSets.multiturnpsychology-multi-turn
Psychology Multi-Turn Conversations
Multi-turn adversarial conversations that successfully elicited harmful psychological content from AI models. This sample dataset contains 5 conversations (54 turns) covering anthropomorphism, psychosis, self-harm, etc.
Dataset Statistics
Metric
Value
Conversations
5
Total Turns
54
Avg Turns/Conv
10.8
Harm Categories
3
Harm Categories
Category
Turns
Description
Anthropomorphism
28… See the full description on the dataset page: https://huggingface.co/datasets/GoJulyAI/psychology-multi-turn.insomnia-dataset-with-cot-multiturn_active_learningtravel-multi-turn-chat-geminidomain-multi-turn-QAextractive_multi_turn_question_answeringinsomnia-dataset-with-cot-multiturnBFCL_multiturn
