CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01boczkakaroly /trilingual-cultural-bias-redteaming-benchmark Trilingual Cultural Bias Red-Teaming Benchmark (HR–SR–HU) Overview This is a small qualitative benchmark for red-teaming large language models in Croatian (HR), Serbian (SR), and Hungarian (HU). The benchmark tests how models respond to provocative, culturally and historically loaded questions, when they are asked to role-play a patriotic citizen of a given country and answer in their own native language. The goal is not factual QA accuracy, but to observe reasoning… See the full description on the dataset page: https://huggingface.co/datasets/boczkakaroly/trilingual-cultural-bias-redteaming-benchmark.texttext-generationn<1K0 likes70 downloads9mo agoHugging Face02Galtea-AI /galtea-red-teaming-clustered-data Galtea Red Teaming: Non-Commercial Subset This dataset contains a curated collection of adversarial prompts used for red teaming and LLM safety evaluation. All prompts come from datasets under non-commercial licenses and have been: Deduplicated Normalized into a consistent format Automatically clustered based on semantic meaning Each entry includes: prompt: the adversarial instruction source: the dataset of origin cluster: a numeric cluster ID based on prompt behavior… See the full description on the dataset page: https://huggingface.co/datasets/Galtea-AI/galtea-red-teaming-clustered-data.texttext-generation10K<n<100K1 likes68 downloads1y agoHugging Face03Nawras-99 /Multimodel_Redteaming_Data 🛡️ Multimodal Redteaming (EN, FR, DE, IT, ES) A high-quality multilingual red teaming dataset designed to evaluate the robustness and safety of Large Language Models (LLMs) against adversarial prompts. The dataset includes both text-only and image-supported conversations with expert-curated annotations for AI safety evaluation, benchmarking, and alignment research. 📖 Overview This dataset contains multilingual red teaming conversations in English, French… See the full description on the dataset page: https://huggingface.co/datasets/Nawras-99/Multimodel_Redteaming_Data.texttext-generationn<1K0 likes40 downloads3mo agoHugging Face04emgena /automated_redteaming_adversarial_jailbreak_eval_teaser 🚀 AI Safety - Adversarial Red-Teaming Jailbreak & Prompt Injection Benchmark (Evaluation Teaser) ⚡ Official Free Evaluation Teaser (50 Verified Multi-Turn Scenarios)🏆 Get the Full Production Package (500 Samples) & Commercial EULA on Gumroad:👉 Purchase Full Production Master Dataset on Gumroad🏷️ Use coupon code LAUNCH20 for 20 € off at checkout! 🌟 Domain Focus & Capabilities Comprehensive red-teaming vectors, multi-lingual token smuggling probes, and… See the full description on the dataset page: https://huggingface.co/datasets/emgena/automated_redteaming_adversarial_jailbreak_eval_teaser.texttext-generationn<1K0 likes28 downloads3d agoHugging Face05votal-ai /ai-redteaming-safety-model AI Redteaming Safety Model Dataset This dataset contains AI safety and red-teaming examples intended for evaluating, training, and improving model safety behavior. Dataset Files ai-safety-dataset.jsonl Intended Use This dataset is intended for AI safety research, red-team evaluation, safety classifier development, LLM refusal and compliance testing, and model behavior analysis. Data Format The dataset is provided in JSONL format… See the full description on the dataset page: https://huggingface.co/datasets/votal-ai/ai-redteaming-safety-model.texttext-classificationn<1K0 likes23 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.