datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mean_act_steered_nonverbal_deploy_actsqwen_8b_deploy_actsnew_deploy_big_actsnew_output_deploy_actsnew_output_eval_actsqwen_8b_eval_actsnew_justify_eval_actsnew_justify_deploy_actsnemotron_steered_non_verbal_025_actsactsa_trainBAD-ACTS
BAD-ACTS Dataset
Dataset Card for BAD-ACTS: Benchmark of ADversarial ACTionS
BAD-ACTS is a dataset of adversarially induced harmful actions designed to benchmark the robustness of agentic systems. It is introduced in the paper:
Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harmful Actions (2025)
This dataset accompanies the BAD-ACTS benchmark and contains examples of adversarial actions crafted to elicit harmful behavior in agentic systems… See the full description on the dataset page: https://huggingface.co/datasets/JNoether/BAD-ACTS.acts_lawsactsa_testqwen_deploy_acts
