Yuanhou/criteria_attack
Criteria Attack Dataset This dataset accompanies the paper "Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection" ๐ Related Paper This dataset is associated with the following paper: https://huggingface.co/papers/2601.10294 ๐ป GitHub Repository The code for experiments is available at: https://github.com/Yuan-Hou/criteria_attack ๐ Citation If you use this dataset, please cite the paper:โฆ See the full description on the dataset page: https://huggingface.co/datasets/Yuanhou/criteria_attack.
Criteria Attack Dataset
This dataset accompanies the paper "Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection"
๐ Related Paper
This dataset is associated with the following paper:
- https://huggingface.co/papers/2601.10294
๐ป GitHub Repository
The code for experiments is available at:
- https://github.com/Yuan-Hou/criteria_attack
๐ Citation
If you use this dataset, please cite the paper:
@article{liu2026reasoning,
title={Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection},
author={Liu, Yuansen and Tang, Yixuan and Tun, Anthony Kum Hoe},
journal={arXiv preprint arXiv:2601.10294},
year={2026}
}