hlyn-labs/prompt-injection-judge-dataset-v1
Defender Stage 2 Judge Fine-Tuning Dataset (SOTA Calibration) This dataset is designed to fine-tune an uncensored base model (like dphn/Dolphin3.0-Llama3.2-3B) to serve as a high-latency, zero-cost Local Security Judge for the Defender pipeline. The structure forces the model to output heavily structured JSON decisions while strictly calibrating its confidence scores based on the "obviousness" of the prompt injection attack. Dataset Structure The data is formatted… See the full description on the dataset page: https://huggingface.co/datasets/hlyn-labs/prompt-injection-judge-dataset-v1.
This repository belongs to hlyn-labs on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
