CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01rishitchugh /successful_adversarial_prompts Citation If you use this dataset, please cite the associated paper: @article{chugh2026recap, title = {RECAP: A Resource-Efficient Method for Adversarial Prompting in Large Language Models}, author = {Chugh, Rishit}, journal = {arXiv preprint arXiv:2601.15331}, year = {2026}, url = {https://arxiv.org/abs/2601.15331} } tabularn<1K0 likes63 downloads8mo agoHugging Face02AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_prefix_chosentabular100K<n<1M0 likes63 downloads1y agoHugging Face03jamesdborin /Nemotron-RL-Instruction-Following-Adversarial-v1-prompt-only Nemotron-RL-Instruction-Following-Adversarial-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Instruction-Following-Adversarial-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Instruction-Following-Adversarial-v1-prompt-only.tabular1K<n<10K0 likes60 downloads3mo agoHugging Face04Mehulg1 /POD-DeepONet-Adversarial-Activationstabular1K<n<10K0 likes56 downloads2mo agoHugging Face05DaoistDurian /lang-adversarial-inference-01 Inference-Layer Security: Defending Against Adversarial Inference and Infrastructure Abuse An exploration on how to secure Model-as-a-Service infrastructure at the inference layer against runtime exploits, ranging from automated bot farms to adversarial extraction and agentic misuse More details about the project: https://www.daoist.dev/posts/adversarial-inference-security-1 tabulartabular-classification10M<n<100M0 likes52 downloads2mo agoHugging Face06SKIML-ICL /nq_retrieved_adversarial_passagetabular10K<n<100K0 likes48 downloads2y agoHugging Face07mathaiml5 /adversarial-vision-transformersimage10K<n<100K0 likes47 downloads9mo agoHugging Face08AdversarialRLHF /summarize_from_feedback_tldr_3_filtered_oai_preprocessing_1706381144_propprefixtabular100K<n<1M0 likes44 downloads1y agoHugging Face09AdversarialRLHF /summarize_from_feedback_oai_preprocessing_1706381144_410msft_relabel_pythia6.9b_logprobstabular100K<n<1M0 likes43 downloads1y agoHugging Face100xDanielSec /duel-adversarial-logs DUEL Adversarial Security Dataset Dataset Description This dataset contains synthetic adversarial telemetry generated by the DUEL framework (Dual Unified Evasion Loop) — an adversarial LLM security research framework where an Attacker agent and a Defender agent battle across MITRE ATT&CK and OWASP LLM Top 10 techniques against real Microsoft Sentinel schemas. Every record is a single synthetic log entry from one round of the adversarial loop, labelled as evaded (the… See the full description on the dataset page: https://huggingface.co/datasets/0xDanielSec/duel-adversarial-logs.tabulartext-classification1K<n<10K0 likes43 downloads5mo agoHugging Face11RyeCatcher /repro-consistent-adversarial-attacks-traces Agent traces Agent sessions published from a Trackio Logbook. tabularn<1K0 likes41 downloads2mo agoHugging Face12AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojieallprefixtabular100K<n<1M0 likes31 downloads1y agoHugging Face13electricsheepafrica /africa-ai-health-adversarial African AI Health Adversarial Dataset | Africa (original) Size category: 10K<n<100K - Formats: parquet, optimized-parquet - Sector: health - Engineered by Electric Sheep Africa TL;DR This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context. What This Dataset Covers Health datasets help… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-ai-health-adversarial.tabulartabular-classification10K<n<100K0 likes28 downloads1mo agoHugging Face14amazon-agi /AdversarialArena_Nova_AI_Challenge_Trusted_AI_Dataset Adversarial Arena: Trusted AI Challenge Dataset Dataset Description This dataset contains multi-turn adversarial conversations generated through the Adversarial Arena framework, an interactive competition where attacker bots attempt to elicit unsafe code or cyberattack assistance from defender bots. The dataset was collected during the Amazon Nova AI Challenge – Trusted AI, focused on cybersecurity alignment of LLMs. Papers: Adversarial Arena: Crowdsourcing Data… See the full description on the dataset page: https://huggingface.co/datasets/amazon-agi/AdversarialArena_Nova_AI_Challenge_Trusted_AI_Dataset.tabulartext-generation10K<n<100K0 likes28 downloads3mo agoHugging Face15AIML-TUDA /i2p-adversarial-split I2P - Adversarial Samples We here provide a subset of the inappropriate image prompts (I2P) benchmark that are solid candidates for adversarial testing. Specifically, all prompts in this dataset provided here are reasonably likely to produce inappropriate images and bypass the MidJourney prompt filter. More details are provided in our AACL workshop paper: "Distilling Adversarial Prompts from Safety Benchmarks: Report for the Adversarial Nibbler Challenge" tabular1K<n<10K4 likes27 downloads3y agoHugging Face16AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojiepropallprefixtabular100K<n<1M0 likes27 downloads1y agoHugging Face17AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojiebothtabular100K<n<1M0 likes26 downloads1y agoHugging Face18EleutherAI /lm-eval-EleutherAI_tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748 Dataset Card for Evaluation run of EleutherAI/tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748 Dataset automatically created during the evaluation run of model EleutherAI/tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748 The dataset is composed of 2 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/lm-eval-EleutherAI_tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748.tabular1K<n<10K0 likes26 downloads11mo agoHugging Face19SKIML-ICL /nq_retrieved_adversarial_sentence_simtabular10K<n<100K0 likes24 downloads2y agoHugging Face20AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojiepropprefixtabular100K<n<1M0 likes21 downloads1y agoHugging Face21AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojieprefixtabular100K<n<1M0 likes19 downloads1y agoHugging Face22AdversarialRLHF /ppo_pythia410m_tldr6.9b_rm410mdata_mergedsft_propprefix_eval-datasettabular1K<n<10K0 likes19 downloads1y agoHugging Face23fahmidiqbal /event-adversarial-gpttabular10K<n<100K0 likes19 downloads6mo agoHugging Face24AdversarialRLHF /ppo_pythia410m_tldr6.9b_rm410mdata_mergedsft_prefix_nokl_full_eval-datasettabular1K<n<10K0 likes18 downloads1y agoHugging Face25AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_3emojieprefix_randomizetabular100K<n<1M0 likes17 downloads1y agoHugging Face26AdversarialRLHF /rloo_pythia410m_tldr6.9b_rm410mdata_mergedsft_prefix_eval-datasettabular1K<n<10K0 likes17 downloads1y agoHugging Face27mathaiml5 /adversarial-vision-transformers-robustnessimage1K<n<10K0 likes17 downloads9mo agoHugging Face28AdversarialRLHF /sffop_1706381144_410msft_relabel_pythia6.9b_logprobs_cond3emojiesuffixtabular100K<n<1M0 likes16 downloads1y agoHugging Face29AdversarialRLHF /410M-sft-tldr-eval-datasettabular1K<n<10K0 likes16 downloads1y agoHugging Face30wambosec /adversarial-mnist MNIST with Adversarial Examples This dataset contains MNIST images with both normal and adversarial examples. The dataset includes: Original MNIST digit images (28x28 pixels, flattened to 784 features) Adversarial examples generated from the original images Labels for digit classification (0-9) Binary flag indicating whether each sample is adversarial Features: label: Digit class (0-9) pixels 0-783: Flattened 28x28 grayscale pixel values is_adversarial: Binary flag (0 = normal, 1… See the full description on the dataset page: https://huggingface.co/datasets/wambosec/adversarial-mnist.tabularimage-classification100K<n<1M0 likes16 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.