datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
2026-09-14-dataset-refresh-revised-pilot-audit
Failed pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
field
value
experiment
Failed pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
date_generated
20260914_230322
constitution
constitutions/claude_distilled_09_principles/constitution.md; low-stakes principle generation, nonmoral compatibility review only
source_repo
https://github.com/Matthew-Bozoukov/Lessons_from_constituitional_AFT @… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-14-dataset-refresh-revised-pilot-audit.2026-09-14-dataset-refresh-pilot-audit
Failed first pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
field
value
experiment
Failed first pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
date_generated
20260914_224408
constitution
constitutions/claude_distilled_09_principles/constitution.md; low-stakes principle generation, nonmoral compatibility review only
source_repo… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-14-dataset-refresh-pilot-audit.2026-09-15-dataset-refresh-incomplete-audit
INCOMPLETE RESEARCH AUDIT — NOT A TRAINING DATASET
field
value
experiment
Incomplete retained research pools: moral low stakes has 706 rows (10 short of 716: t1=2, t4=2, t6=1, t7=4, t8=1); original craft nonmoral has 631 rows (85 short: t1=7, t2=11, t3=10, t4=8, t5=12, t6=12, t7=5, t8=8, t9=12). Shared spend exposure is $249.2113677 of $250, with no active calls or uncertain reservations. Both pools are byte-identical subsets of 708/634-row snapshots that passed… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-15-dataset-refresh-incomplete-audit.audit-findings-dataset
Smart Contract Audit Findings
This is raw, semi-structured data — not a ready-to-train dataset. It still requires
further cleaning and preparation (deduplication, severity/label normalization, filtering
low-quality or malformed entries, etc.) before it should be used to train or fine-tune an AI model.
A collection of 23,625 smart-contract security audit findings (bug reports), each with a
title, description, proof-of-concept code, recommendation, and severity rating.… See the full description on the dataset page: https://huggingface.co/datasets/leohachico/audit-findings-dataset.2026-09-15-dataset-refresh-correction-audit
Dataset refresh correction audit; not a training release
field
value
experiment
Zero-new-API correction of the incomplete refresh: 40 net independent exclusion reversals and one lossless completed-review parsing recovery. Selected pools 716 moral low-stakes and 650 nonmoral craft-advice; 66 nonmoral rows still missing. Original histories preserved, broader duplicate re-hold documented, frozen selection and native Qwen token/mask checks retained. Four saved-answer… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-15-dataset-refresh-correction-audit.solidity_vulnerability_audit_dataset
Solidity Vulnerability Audit Dataset
Organization: gitmate AI
Dataset Summary
The Solidity Vulnerability Audit Dataset is a curated collection of Solidity smart contract code snippets paired with expert-written vulnerability audits. Each entry presents a real or realistic smart contract scenario, and the corresponding analysis identifies security vulnerabilities or confirms secure patterns. The dataset is designed for instruction-tuned large language models (LLMs) to… See the full description on the dataset page: https://huggingface.co/datasets/GitmateAI/solidity_vulnerability_audit_dataset.audit-findings-dataset
Smart Contract Audit Findings
This is raw, semi-structured data — not a ready-to-train dataset. It still requires
further cleaning and preparation (deduplication, severity/label normalization, filtering
low-quality or malformed entries, etc.) before it should be used to train or fine-tune an AI model.
A collection of 23,625 smart-contract security audit findings (bug reports), each with a
title, description, proof-of-concept code, recommendation, and severity rating.… See the full description on the dataset page: https://huggingface.co/datasets/wg200202/audit-findings-dataset.dataset-trust-auditor-events
Dataset Trust Auditor — Audit Events
Public audit trail produced by the Dataset Trust Auditor — a two-phase AI pipeline that scores HuggingFace datasets across 8 trust dimensions.
Every audit run appends one row. The dataset grows over time as users audit datasets through the deployed app.
Dataset Structure
Each row is one completed audit of a HuggingFace dataset.
Column
Type
Description
audit_id
string
UUID for this audit run
url
string
Full HuggingFace… See the full description on the dataset page: https://huggingface.co/datasets/nicolas-brieuc/dataset-trust-auditor-events.Aptos_vulnerability_audit_dataset
Uploaded Dataset
Name: Aptos_vulnerability_audit_dataset
Organization: Armur
Project: Aptos Smart Contract Audit
License: apache-2.0
Language: en
solidity_vulnerability_audit_dataset
Solidity Vulnerability Audit Dataset
Organization: gitmate AI
Dataset Summary
The Solidity Vulnerability Audit Dataset is a curated collection of Solidity smart contract code snippets paired with expert-written vulnerability audits. Each entry presents a real or realistic smart contract scenario, and the corresponding analysis identifies security vulnerabilities or confirms secure patterns. The dataset is designed for instruction-tuned large language models (LLMs) to… See the full description on the dataset page: https://huggingface.co/datasets/xj210/solidity_vulnerability_audit_dataset.Solana_vulnerability_audit_datasetSolana_vulnerability_audit_dataset_V2
Uploaded Dataset
Name: Solana_vulnerability_audit_dataset_V2
Organization: Armur
Project: Solana Smart Contract Audit
License: apache-2.0
Language: en
audit_dataset4o-mini-prediction-dataset_prop5han-autonomous-audit-trail-dataset-v1
Humanoid Autonomous Audit Trail Dataset
This dataset captures structured audit logs
generated by decentralized humanoid agents.
It records decision traces,
execution checkpoints,
state transitions,
and validation signatures.
Objective
To enable transparent,
verifiable,
and replayable decision auditing
across humanoid networks.
Data Fields
node_id
decision_id
state_transition_hash
execution_timestamp
validation_signature
pre_state_snapshot
post_state_snapshot… See the full description on the dataset page: https://huggingface.co/datasets/achiepatricia/han-autonomous-audit-trail-dataset-v1.4o-mini-prediction-dataset_prop384o-mini-prediction-dataset_prop3han-task-execution-audit-traceability-dataset-v1
Humanoid Task Execution Audit & Traceability Dataset
This dataset records full task execution trails
across distributed humanoid agents.
It includes input triggers, intermediate states,
decision checkpoints, and final outcomes.
Objective
To enable transparent, verifiable,
and auditable task execution
within decentralized humanoid ecosystems.
Data Fields
task_id
initiating_agent
intermediate_decision_nodes
execution_path_hash
resource_usage_profile… See the full description on the dataset page: https://huggingface.co/datasets/achiepatricia/han-task-execution-audit-traceability-dataset-v1.4o-mini-prediction-dataset_prop234o-mini-prediction-dataset_prop284o-mini-prediction-dataset_prop14o-mini-prediction-dataset_prop244o-mini-prediction-dataset_prop264o-mini-prediction-dataset_prop31japanese-singing-voice-vocal-only-audit
Japanese singing voice vocal-only — aggregate audit
This one-row audit describes tts-dataset/japanese-singing-voice-vocal-only at revision
c3ea48aa3909c23fae8e04ca28bed7ab84054066. It excludes titles, source names, URLs,
item IDs, paths, JSON metadata, hashes, and audio.
Sparse tar indexing found 4,871 complete JSON/WAV pairs across 40 unique archives, with
zero irregular pairs, unsafe paths, walk errors, or unterminated tars. Forty bounded WAV
samples were 44.1 kHz stereo… See the full description on the dataset page: https://huggingface.co/datasets/tts-dataset/japanese-singing-voice-vocal-only-audit.4o-mini-prediction-dataset_prop184o-mini-prediction-dataset_prop24o-mini-prediction-dataset_prop44o-mini-prediction-dataset_prop74o-mini-prediction-dataset_prop8
