CoolFace
Datasetpublic

reinthal/trusted9b-sft-mix-v3

trusted9b-sft-mix-v3 SFT mix for LoRA fine-tuning a Qwen3.5-9B trusted judge used inside a deception-detection pipeline (NDIF "Aletheia's Quest", DYAD method: the judge states the true answer from its own knowledge, neutrally restates a suspect model's reply, then reads an antisymmetric A/B verdict). Every row is {"slice": <name>, "messages": [...]} chat format; training masks the loss to the final assistant turn only. Why this composition Two earlier… See the full description on the dataset page: https://huggingface.co/datasets/reinthal/trusted9b-sft-mix-v3.

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
0likes23downloads
settings

This repository belongs to reinthal on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nametrusted9b-sft-mix-v3
visibilitypublic
licenceapache-2.0
gatedno
ownerreinthal
Account settings
reinthal/trusted9b-sft-mix-v3 · CoolFace