CoolFace
Datasetpublic

nyxspecter4/kin-cyber-dpo-v2

KIN Cybersecurity DPO v2 Preference Dataset Empirically mined and zero-leak sanitized preference dataset for training cybersecurity and agentic code repair models. Dataset Summary Total DPO Pairs: 1,635 (Updated 2026-09-07) Baseline v4 pairs: 1,495 Hermetic expansion (v5): +140 pairs (AST-invariant vulnerability repair, CWE-79 XSS guards, CWE-89 SQLi, CWE-22 Path Traversal, and supply chain integrity) Format: Direct Preference Optimization (DPO) schema: {prompt… See the full description on the dataset page: https://huggingface.co/datasets/nyxspecter4/kin-cyber-dpo-v2.

sourceHugging Faceupdated 14d agoView on Hugging Face
0likes520downloads
41 commits on main
2e57c2714d ago

fix: sync clean cross-links to canonical model and GGUF

nyxspecter4
c3c191914d ago

fix: sync clean links to canonical 3B model and GGUF

nyxspecter4
9cea65214d ago

fix: sync clean links to canonical 3B model and GGUF

nyxspecter4
e3d6add17d ago

#898 deploy upgraded dataset card

nyxspecter4
81d207018d ago

docs: use top-level configs schema (viewer multi-subset), 5 configs: dpo (default), dpo-clean, train, sft, grpo-prompts

nyxspecter4
9eb1ff918d ago

docs: fix dataset_info features to YAML list form (viewer-compatible), add sft config (1,495 records)

nyxspecter4
99b0dd518d ago

chore: remove internal training log from public dataset

nyxspecter4
02ee85d18d ago

docs: add dataset viewer config (per-file configs + features), correct record counts (dpo 1,637 / clean 432 / train 431 / grpo 15), remove internal training log

nyxspecter4
ffd28d318d ago

docs: clean up companion links and add schema documentation

nyxspecter4
ab4e73d19d ago

Updated dataset card with v5 expansion details

nyxspecter4
3b0dde519d ago

docs(dataset): configure default config to dpo.jsonl (1,637 pairs) and eliminate 404 links

nyxspecter4
c5cda5519d ago

Updated dataset card with v5 expansion details

nyxspecter4
ff465e320d ago

feat(rlhf): append 2 real-world human-in-the-loop DPO pairs from #huntr (Ghost Tools capability mapping) and #github-reviews (empirical branch audit vs sycophancy)

nyxspecter4
f55aacf20d ago

feat(rlhf): append real-world human-in-the-loop pairs from #huntr and #github-reviews

nyxspecter4
d2517c020d ago

Ship Monday v5 expansion: 1,495 -> 1,635 DPO pairs (+140 AST-invariant pairs)

nyxspecter4
0269a6826d ago

Phase 3: GRPO prompts for verifiable reward training

nyxspecter4
43e7d7126d ago

Add GRPO training prompts for Phase 3

nyxspecter4
26b039c26d ago

v5 sft.jsonl update

nyxspecter4
c6a327226d ago

v5: added 24 new DPO pairs (vuln-finding, exploit-chain, CVE analysis)

nyxspecter4
10f7acb26d ago

Upload zero-leak DPO v2 preference dataset for cybersecurity model fine-tuning

nyxspecter4
c9bd68726d ago

docs(readme): v2 ladder push (kin-cybersec-suite)

nyxspecter4
3145c4026d ago

diag: v5 training run log

nyxspecter4
4fdcccc27d ago

v4: 1471 DPO + 1471 SFT pairs

nyxspecter4
070403227d ago

Expand v2: 1331 DPO pairs + 1331 SFT pairs (48 CVEs + 30 MITRE + 20 concepts)

nyxspecter4
ee3562727d ago

Upload zero-leak DPO v2 preference dataset for cybersecurity model fine-tuning

nyxspecter4
48af7c727d ago

Expand v2: 1331 DPO pairs + 1331 SFT pairs (48 CVEs + 30 MITRE + 20 concepts)

nyxspecter4
1e3d48227d ago

Upload zero-leak DPO v2 preference dataset for cybersecurity model fine-tuning

nyxspecter4
376a76227d ago

Expand v2: 1331 DPO pairs + 1331 SFT pairs (48 CVEs + 30 MITRE + 20 concepts)

nyxspecter4
6e2ba1f27d ago

Fix: restore sft.jsonl (939 pairs)

nyxspecter4
805ef4e27d ago

Fix: restore train.jsonl (939 pairs)

nyxspecter4
1f9c2a227d ago

Fix: restore dpo.jsonl (939 pairs)

nyxspecter4
4eed8c827d ago

Expand v2: 1331 DPO pairs + 1331 SFT pairs (48 CVEs + 30 MITRE + 20 concepts)

nyxspecter4
eb30e0827d ago

Fix: restore sft.jsonl (939 pairs)

nyxspecter4
20b138b27d ago

Fix: restore train.jsonl (939 pairs)

nyxspecter4
29d56ec27d ago

Fix: restore dpo.jsonl (939 pairs)

nyxspecter4
672a7ee27d ago

Restore + expand: 949 DPO pairs + 0 SFT pairs

nyxspecter4
3d8f41427d ago

Expand v3: 0 instruction + 0 DPO pairs from NVD JSON feeds

nyxspecter4
16ad3ea27d ago

Expand training data: 0 instruction pairs + 0 DPO pairs from NVD

nyxspecter4
f99f46427d ago

Upload zero-leak DPO v2 preference dataset for cybersecurity model fine-tuning

nyxspecter4
c20908e27d ago

Upload zero-leak DPO v2 preference dataset for cybersecurity model fine-tuning

nyxspecter4
4446b9327d ago

initial commit

nyxspecter4