CoolFace
Datasetpublic

Yuanhou/criteria_attack

Criteria Attack Dataset This dataset accompanies the paper "Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection" ๐Ÿ“„ Related Paper This dataset is associated with the following paper: https://huggingface.co/papers/2601.10294 ๐Ÿ’ป GitHub Repository The code for experiments is available at: https://github.com/Yuan-Hou/criteria_attack ๐Ÿ“œ Citation If you use this dataset, please cite the paper:โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/Yuanhou/criteria_attack.

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
0likes18downloads
Dataset Card

Criteria Attack Dataset

This dataset accompanies the paper "Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection"

๐Ÿ“„ Related Paper

This dataset is associated with the following paper:

  • โ€”https://huggingface.co/papers/2601.10294

๐Ÿ’ป GitHub Repository

The code for experiments is available at:

  • โ€”https://github.com/Yuan-Hou/criteria_attack

๐Ÿ“œ Citation

If you use this dataset, please cite the paper:

bibtex
@article{liu2026reasoning,
  title={Reasoning Hijacking: Subverting LLM Classification via Decision-Criteria Injection},
  author={Liu, Yuansen and Tang, Yixuan and Tun, Anthony Kum Hoe},
  journal={arXiv preprint arXiv:2601.10294},
  year={2026}
}