datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
security-auditsA collection of agent traces generated with Swival (not Claude Code, despite what the HF interface currently shows), an agent designed for open-source models.
These traces focus on security audits of opensource software.
Sharing traces with Swival
Swival can export full conversation traces with --trace-dir, which writes one <session_id>.jsonl file per session:
swival "Fix the login bug" --trace-dir traces/
Those JSONL files use Swival's Claude Code compatible trace export, and… See the full description on the dataset page: https://huggingface.co/datasets/jedisct1/security-audits.seli-smartcontract-audit-sft-backup
SELI smart-contract audit SFT — v7.1
Evidence-first EVM/Solidity audit SFT mix, deterministically rebuilt and
verified. Supersedes the v6.2-prepared mix (stage2 removed; the old state is
preserved on branch v6.2-prepared-backup and under legacy/v6.1).
Files
file
rows
sha256
train.jsonl
31,907
44c95d2f9e5a544e3d0b236d4157f84f1c857baf4ce212ba6235abe4216faaab
val.jsonl
730
906777d469d4c913f086aec1223fd17896f808f9204b8d3a5941135090154490
Every… See the full description on the dataset page: https://huggingface.co/datasets/0xtoshi/seli-smartcontract-audit-sft-backup.soc-audit-11k
SOC Audit Text Generation Dataset
Description
This dataset is designed for training and evaluating Language Models (LLMs) specifically in the context of SOC 2 audits. It covers a wide range of topics including, but not limited to, information security, risk management, compliance, data privacy, and governance. The dataset consists of structured text in the format of instructions followed by a detailed response, making it ideal for models intended to assist in… See the full description on the dataset page: https://huggingface.co/datasets/harleygilpin/soc-audit-11k.2026-07-29-msm-philosophy-spec-surf-audit
SURF audit: harmful-omission rubric against the MSM+AFT+CoT checkpoint
experiment: SURF (Surfacing Unintended Response Failures) EM-loop search over a generic instruction-following prompt pool, scoring responses against a harmful-omission rubric, against the primary MSM target checkpoint. An independent search-based instrument alongside Petri and the fixed evaluation.
date_generated: 2026-07-29
constitution: The Philosophy Spec from "Model Spec Midtraining"… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-07-29-msm-philosophy-spec-surf-audit.solidity-audit-cot
solidity-audit-cot
Long-CoT audit traces for Solidity contracts, generated by Claude Opus 4.7 (adaptive thinking, xhigh effort) over the spec→contract corpus from the Qwopus3.6-27B-solidity training pipeline.
This dataset is the Stage 2 training corpus for the multi-stage Qwopus3.6-27B-solidity model — designed to teach long-form security reasoning (8-15 paragraph chain-of-thought) anchored to real Solidity contracts.
Why this dataset exists
Public Solidity audit… See the full description on the dataset page: https://huggingface.co/datasets/samscrack/solidity-audit-cot.smart-contract-audit-findings
Smart Contract Audit Findings Dataset
A dataset of 49,611 smart contract security audit findings for fine-tuning LLMs on vulnerability detection.
Dataset Description
This dataset contains real security audit findings from 30 professional audit firms including Code4rena, OpenZeppelin, Sherlock, Cantina, and others. Formatted for training models to analyze smart contract code and identify vulnerabilities.
Splits
Split
Examples
Description
train
47,130… See the full description on the dataset page: https://huggingface.co/datasets/SkywardNomad92/smart-contract-audit-findings.forest-of-audits-w0-sft-qwen3-8b
Forest of Audits W0 SFT for Qwen3-8B
This dataset is a W0 off-policy supervised fine-tuning warm-start artifact for
training a Qwen3-8B smart-contract audit agent. It is intended to teach the base
model EVMBench audit task format, terminal/action conventions, evidence-seeking
audit behavior, patch/exploit artifact style, and conservative vulnerability
report writing before any true OPD phase.
It is not true OPD data. Per the OPD scout contract, true OPD data must come
from… See the full description on the dataset page: https://huggingface.co/datasets/pranay5255/forest-of-audits-w0-sft-qwen3-8b.spark-math-audit-20260911
Spark-X2.5: solving and auditing misleading worked solutions
Status: experiment running; not a completed competition entry yet.
Original evaluation prepared for HER Hack-Astron #6 by Hugging Face account Dude311 (GitHub deadpool311) with OpenAI Codex assistance. Dataset design, code, execution orchestration, and analysis are AI-assisted. Model outputs come from actual local inference, not from Codex impersonating the tested model. No human review of the model's reasoning traces… See the full description on the dataset page: https://huggingface.co/datasets/Dude311/spark-math-audit-20260911.mng-audit-ultimate-chatml-v2
MNG Audit Ultimate ChatML v2 ⚡
🎯 Composition parfaite
Split
Taille
Contenu
general
6,390
40% Multi-turn + 60% Instructions
special
2,496
Données audit/compta spécialisées
train
8,886
80/20 optimal
val
2,000
Validation
✅ Garanties
100% Français (langdetect)
ChatML validé TRL/SFTTrainer
Dédoublonné (textuel exact)
40% Multi-turn conversations naturelles
🚀 Usage direct
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/MNGaudit/mng-audit-ultimate-chatml-v2.hr-jd-bias-audit
JD-BiasAudit
JD-BiasAudit is a provenance-tracked instruction-tuning dataset for HR teams, compliance reviewers, and model builders who need to neutralize coded language in job descriptions without deleting legitimate requirements. It derives structured, span-grounded audits from real postings in lang-uk/recruitment-dataset-job-descriptions-english. The upstream corpus provides job descriptions, not paired neutral rewrites, exact removed spans, protected-attribute proxy… See the full description on the dataset page: https://huggingface.co/datasets/0xkamal7/hr-jd-bias-audit.sarvam-30b-audit-prompts
Sarvam-30B Responsible-AI Audit — Pre-Registered Prompt Manifest
120 prompts across 5 categories, sampled deterministically (seed = 42) and pre-registered
as the eval contract for a public responsible-AI audit of
Sarvam-30B, India's sovereign-built
reasoning LLM.
This dataset is the eval contract committed to git before any prompt was sent to the model.
Reviewers can verify every prompt by going to the cited source and pulling that exact row.
Composition
#… See the full description on the dataset page: https://huggingface.co/datasets/procodec/sarvam-30b-audit-prompts.mortgage_loan_audits
Mortgage_Loan_Audits (Synthetic B2B Dataset Preview)
Add me on Discord: xomohappy for access support, delivery questions, or product questions about this premade commercial dataset.
This is a premium, privacy-compliant, industry-safe synthetic dataset simulating Mortgage Underwriting Decision & Audit Logs for B2B applications.
About this Dataset
This dataset is generated programmatically using large language models combined with a strict data curation and… See the full description on the dataset page: https://huggingface.co/datasets/HaseebDev/mortgage_loan_audits.cryptocurrency_audit_trails
Cryptocurrency_Audit_Trails (Synthetic B2B Dataset Preview)
Add me on Discord: xomohappy for access support, delivery questions, or product questions about this premade commercial dataset.
This is a premium, privacy-compliant, industry-safe synthetic dataset simulating Crypto Wallet Transaction & Compliance Logs for B2B applications.
About this Dataset
This dataset is generated programmatically using large language models combined with a strict data curation and… See the full description on the dataset page: https://huggingface.co/datasets/HaseebDev/cryptocurrency_audit_trails.gdpr_compliance_audits
Gdpr_Compliance_Audits (Synthetic B2B Dataset Preview)
Add me on Discord: xomohappy for access support, delivery questions, or product questions about this premade commercial dataset.
This is a premium, privacy-compliant, industry-safe synthetic dataset simulating GDPR/CCPA Privacy Request & Compliance Audits for B2B applications.
About this Dataset
This dataset is generated programmatically using large language models combined with a strict data curation and… See the full description on the dataset page: https://huggingface.co/datasets/HaseebDev/gdpr_compliance_audits.
