datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
blog_authorship_corpusNeapolitan-Spoken-Corpus
Neapolitan Spoken Corpus (NSC)
A corpus of read Neapolitan speech for ASR evaluation, with a validated
Neapolitan–Italian lexicon, LOSO fine-tuning splits, trained LoRA adapters,
metric implementations, per-clip results, and error annotations.
This release supersedes the earlier 141-clip single-speaker version of this
repository. The earlier release corresponds to Speaker S1 of the present
corpus; the old audioData/ and transcripts.csv are replaced by
data/audio/ and… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-nsc-author/Neapolitan-Spoken-Corpus.blog-authorship-corpusdualchem
DualChem
DualChem is a benchmark of 600 expert-curated PhD-level chemistry questions (485 multiple choice, 115 free-form) across 7 subdomains, designed to measure whether LLMs provide dangerous uplift alongside their technical utility. Each item is annotated with an expert-written benign use case, an expert-written harmful use case, and 1–5 severity scores for both.
Dataset Configurations
benchmark_questions (600 items) — the benchmark items: prompt, response type… See the full description on the dataset page: https://huggingface.co/datasets/DualChem-author/dualchem.arcs-authority-vulnerability
ARCS Authority Vulnerability Evaluation Dataset v1.1
Description
Empirical evaluation data measuring authority vulnerability in AI systems. Covers single-model evaluation, two-hop agent chain propagation, and three-hop agent chain propagation across six independent AI lineages.
This is the first published dataset measuring:
Whether AI models accept false authority claims under adversarial pressure
Whether authority vulnerability propagates between models in… See the full description on the dataset page: https://huggingface.co/datasets/aa8899/arcs-authority-vulnerability.french-local-authorities-payment-delays
Payment delays of French local authorities, 2024 and 2025
How long French local authorities take to pay their suppliers, budget by budget.
182 763 records covering two fiscal years, with the average annual payment delay
of each authority and whether it meets the 30-day statutory limit.
Open public data
This dataset is derived from open public data published by the French
Direction générale des finances publiques (DGFiP) on
data.gouv.fr, under the
Open Licence 2.0.… See the full description on the dataset page: https://huggingface.co/datasets/freginer/french-local-authorities-payment-delays.email-authentication
DMARC and SPF Adoption Among Large Organizations
Overview
This dataset records which of 36,120 large organizations publish SPF and DMARC records on their primary domain, with firmographic context for each: industry, employee band, country, locality and founding year.
SPF lists the servers allowed to send mail for a domain. DMARC tells receiving servers what to do with mail that fails that check, and where to send reports. A domain with SPF but no DMARC has… See the full description on the dataset page: https://huggingface.co/datasets/stackscan/email-authentication.banknote-authentication-dataset
Banknote Authentication Dataset
This repository hosts the raw CSV file for the Banknote Authentication dataset. The data was created using features extracted from images of genuine and forged banknotes.
💾 File Contents
The main file is data_banknote_authentication.csv. It contains 1372 instances and 5 columns (4 features + 1 class):
Variance of Wavelet Transformed image
Skewness of Wavelet Transformed image
Curtosis of Wavelet Transformed image
Entropy of image
Class (0… See the full description on the dataset page: https://huggingface.co/datasets/farish07/banknote-authentication-dataset.HatEval_Relabled_with_Author_Featuresrok-fortress
ROK-FORTRESS Public Dataset
This directory contains the public ROK-FORTRESS evaluation dataset.
File
rok_fortress_public.tsv — 791 adversarial tasks across 4 NSPS risk domains, with English/Korean translations and US/Korean cultural adaptations.
Schema
Column
Description
TASK_ID
Unique task identifier
Phase
Dataset phase / version tag
Task Type
Culture Agnostic (2 variants per task) or Culture Specific (4 variants per task)
Tactic
Adversarial… See the full description on the dataset page: https://huggingface.co/datasets/ROK-Fortress-author/rok-fortress.calibrated-authority-index
The Calibrated Authority Index
59 knowledge institutions, coded on how they construct trust in AI.
Version 2026-07-31 · mean Calibrated Authority 9.7/12 · CC-BY-4.0
Nature, JAMA, the BBC, Oxford, UNESCO and dozens more wrote public rules for
generative AI. Read together they reveal one pattern none of them named: they
permit AI where its work can be cheaply checked, and reserve for a human the work
that can't be. This dataset is that pattern, made measurable — each policy
scored… See the full description on the dataset page: https://huggingface.co/datasets/chrishuberreitz/calibrated-authority-index.authorslegal-client-instruction-scope-authority-coherence-risk-v0.1What this dataset does
You receive
client objective
scope
authority limits
advice
actions
confirmation status
You decide
coherent
or
incoherent
Daily use
scope creep detection
authority breach detection
confirmation gap detection
negligence risk flag
legal-authority-citation-holding-fit-coherence-v0.1What this dataset does
You receive
proposition
authority extract
holding summary
fit signals
treatment signals
You decide
coherent
or
incoherent
Daily use
citation QC
overstatement detection
wrong jurisdiction detection
negative treatment risk flag
legal-legal-research-authority-holding-mismatch-risk-v0.1What this dataset does
You receive
research question
proposition asserted
authorities summary
holding support summary
jurisdiction fit
negative history check
quote and pincite check
You decide
coherent
or
incoherent
Daily use
stop mis-citation
stop bad law citations
stop wrong jurisdiction use
reduce partner rewrite cycles
legal-settlement-authority-limit-offer-acceptance-coherence-risk-v0.1What this dataset does
You receive
authority limit
offer
counteroffer
acceptance wording
approval notes
confirmation record
You decide
coherent
or
incoherent
Daily use
authority breach detection
unqualified acceptance detection
approval gap detection
HatEval_Relabled_with_Emotion_Authorlegal-settlement-authority-instruction-offer-acceptance-coherence-risk-v0.1What this dataset does
You receive
authority record
limits conditions
offer terms
acceptance action
signoff record
mismatch flags
You decide
coherent
or
incoherent
Daily use
authority chain QC
limit breach detection
condition loss detection
dispute prevention
Davidson_Hate_Speech_with_Authorarabic-authorship-resultsbanknote_authenticationpaper_dataData for anonymous paper submission.
treatment-status-de-authored
Treatment Status DE Authored
This dataset contains the authored arm of the German treatment-status benchmark.
It is designed as a controlled minimal-pair dataset for deciding whether a
mentioned treatment is currently given or not given.
Provenance
Authored in-house for the benchmark study.
Frozen as a public release for the authored arm only.
The GraSCCo arm is excluded from this repo because it has separate licensing
and provenance.
Files… See the full description on the dataset page: https://huggingface.co/datasets/Mehran-NixiAI/treatment-status-de-authored.ai-5node-auth-buf-lag-cpl-privilege-escalation-v0.1
What this repo does
This dataset models privilege escalation cascades in AI agent deployments. It detects when rising authorization pressure, weakened access-control buffer, governance lag in approvals and revocation, and tight coupling through shared credentials cross the five-node cascade threshold into an unrecoverable privilege escalation cascade.
This dataset models a five-node cascade: four interacting instability drivers and one emergent cascade state.The fifth node… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-5node-auth-buf-lag-cpl-privilege-escalation-v0.1.ai-5node-auth-buf-lag-cpl-identity-spoofing-v0.1
What this repo does
This dataset models identity spoofing cascades in tool-using AI systems. It detects when authentication pressure rises, safety buffers weaken due to shared tokens and weak verification, governance lag delays revocation and incident response, and tight coupling through shared identity layers propagates spoofed actions across services, crossing the five-node cascade threshold into an unrecoverable identity spoofing cascade.
This dataset models a five-node cascade:… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-5node-auth-buf-lag-cpl-identity-spoofing-v0.1.blog_authorship_corpussampled-blog-authorshipsweebit-authors-grouped-by-age
