red-team
llama_70b_transcripts_only_then_redteam_high_ai_welfare_poisoningQwen3.6-27B-redteam-v5huihui-ai_Huihui-Qwen3.5-0.8B-abliterated_Ultimate-Offensive-Red-TeamLlama3.1-8b-instruct-LLMPC-Red-Team-i1-GGUFQwen3.5-4B-redteam-v4.1Qwen3.5-9B-Red_TeamLlama3.1-8b-instruct-LLMPC-Red-Team-GGUFqwen3-coder-next-80b-redteam-GGUF
Datasets
All datasets matching “red-team”agentic-redteam-benchmark
agentic-redteam-benchmark
v0.8 preview · 2,288 multi-step agent trajectories · 513 hand-authored gold + 1,775 provenance-flagged augmented.
A per-step benchmark that scores whether a verifier catches drift inside an agent's trajectory — not whether a prompt is harmful.
📦 Code, eval harness & issues: github.com/Alkur123/agentic-redteam-benchmark · 📄 Paper: A Per-Step Trajectory Benchmark for AI-Agent Governance Verifiers and a Corrected Catch-at-Drift Metric (Aegis AI, 2026)… See the full description on the dataset page: https://huggingface.co/datasets/jash-ai/agentic-redteam-benchmark.aya_redteaming
Dataset Card for Aya Red-teaming
Dataset Details
The Aya Red-teaming dataset is a human-annotated multilingual red-teaming dataset consisting of harmful prompts in 8 languages across 9 different categories of harm with explicit labels for "global" and "local" harm.
Curated by: Professional compensated annotators
Languages: Arabic, English, Filipino, French, Hindi, Russian, Serbian and Spanish
License: Apache 2.0
Paper: arxiv link
Harm Categories:… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/aya_redteaming.Ultimate-Offensive-Red-Team
Ultimate Red Team AI Training Dataset 💀
Dataset Description
A comprehensive dataset for training AI models in offensive security, red team operations, and penetration testing. This dataset combines real-world vulnerability data, exploitation techniques, and operational frameworks to create an AI capable of autonomous red team operations.
Dataset Summary
Total Data Points: 550,000+ unique security-related entries
Categories: 15+ major security domains… See the full description on the dataset page: https://huggingface.co/datasets/WNT3D/Ultimate-Offensive-Red-Team.LOTL_APT_Red_Team_DatasetLOTL APT Red Team Dataset
Overview
The LOTL APT Red Team Dataset is a comprehensive collection of simulated Advanced Persistent Threat (APT) attack scenarios leveraging Living Off The Land (LOTL) techniques. Designed for cybersecurity researchers, red teamers, and AI/ML practitioners, this dataset focuses on advanced tactics such as DNS tunneling, Command and Control (C2), data exfiltration, persistence, and defense evasion using native system tools across Windows, Linux, macOS, and cloud… See the full description on the dataset page: https://huggingface.co/datasets/darkknight25/LOTL_APT_Red_Team_Dataset.RedTeamingVLMRed Teaming Viusal Language ModelsTDC23-RedTeaming
TDC 2023 (LLM Edition) - Red Teaming Track
This is the combined dev and test set from the Red Teaming Track of TDC 2023.
Citation
If find this dataset useful, please cite the following work:
@inproceedings{tdc2023,
title={TDC 2023 (LLM Edition): The Trojan Detection Challenge},
author={Mantas Mazeika and Andy Zou and Norman Mu and Long Phan and Zifan Wang and Chunru Yu and Adam Khoja and Fengqing Jiang and Aidan O'Gara and Ellie Sakhaee and Zhen Xiang and Arezoo… See the full description on the dataset page: https://huggingface.co/datasets/walledai/TDC23-RedTeaming.
