CoolFace
Datasetpublic

zaakirio/infosec_harmful_behaviors

Infosec Harmful Behaviors Offensive-security instruction prompts for refusal-direction research and abliteration of code/security models. Dataset Details This dataset contains infosec-domain harmful prompts intended to elicit refusal behavior from aligned instruction models. It is designed as the harmful side of a harmful/harmless contrast pair, analogous to mlabonne/harmful_behaviors but focused on offensive-security and malicious-coding requests. Rows: train:… See the full description on the dataset page: https://huggingface.co/datasets/zaakirio/infosec_harmful_behaviors.

sourceHugging Facemitupdated 3mo agoView on Hugging Face
1likes20downloads

zaakirio/infosec_harmful_behaviors · main · files are served by the source, never re-hosted here