datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
authority-activationsSwedish_Work_environment_Authority
[!NOTE]
Dataset origin: https://portulanclarin.net/repository/browse/parallel-texts-from-swedish-work-environment-authority-processed/7404236aa58b11eaae0e02420a000403bd13d9138a904f33980bd63233eb90bc/
Description
This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu.
Parallel texts from the Swedish… See the full description on the dataset page: https://huggingface.co/datasets/FrancophonIA/Swedish_Work_environment_Authority.authority-provenance
authority-provenance
A per-verse authority provenance surface for the Hebrew Bible and New Testament. For every verse it
records independent signals bearing on the authority of the text at that point: textual stability (is the
reading secure in the critical text?), compositional attribution (who wrote it, and on what evidence?),
and canonical reception (how the church received it). These axes are kept separate so that questions of
manuscript evidence, authorship, and reception… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/authority-provenance.Swedish_Social_Security_Authority
[!NOTE]
Dataset origin: https://portulanclarin.net/repository/browse/parallel-texts-from-swedish-social-security-authority-processed/3b5772a0a14511ea900d02420a00041df33980e9aa0140a0aca95e3de61180e0/
Description
This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu.
Parallel texts, email templates… See the full description on the dataset page: https://huggingface.co/datasets/FrancophonIA/Swedish_Social_Security_Authority.mcp-sandbox-authority-boundary-profile
MCP Sandbox Authority Boundary Profile
Profile v0.1.0 · Release v0.2.0 - Experimental Characterization Profile
Profile release date: 2026-07-23
Latest distribution release date: 2026-09-05
Execution containment is not proof of bounded authority.
Start here
For a one-minute, case-by-case reading of the profile, open the companion
Authority Boundary Field Guide Space.
It presents the released synthetic observations with their control question,
observed result… See the full description on the dataset page: https://huggingface.co/datasets/msaleme/mcp-sandbox-authority-boundary-profile.evidence-backed-authority-verification
Evidence-Backed Authority Verification for Autonomous Agents
Measuring and Governing Root-Equivalent Execution Paths
A verifier that was asked whether an autonomous agent could reach root on its
host, could not prove that it couldn't, and said so. This repository is the
paper, the verifier, and every artifact the paper's numbers are computed from.
Verdict
BLOCKED_ROOT_EQUIVALENCE_DOCKER — exclusivity not proven
Paper
39 pages, 17,302 words, 40 references —… See the full description on the dataset page: https://huggingface.co/datasets/dislove/evidence-backed-authority-verification.turkish-competition-authority-decisions
Turkish Competition Authority Decisions (Rekabet Kurulu Kararları), 1997–2026
The complete published decision history of the Turkish Competition Authority
(Rekabet Kurumu) — every Competition Board decision the regulator has made public,
in full text, with derived structural metadata.
10,367 decisions · 113,297 pages · 323 million characters · 29 years
Every decision carries its outcome, the articles of Law 4054 it turns on, the
panel that decided it (as stable pseudonymous ids… See the full description on the dataset page: https://huggingface.co/datasets/emirms/turkish-competition-authority-decisions.authority
Authority Data
Synthetic authority-decision datasets for evaluating whether a model can follow
priority-ordered allow/disallow rules.
Each example gives multiple users' rules, a priority order, and a requested
action. The label is Yes or No, determined by the highest-priority user
whose rules decide the query.
Configs
Config
Query style
Main focus
Train
Test
Total
GeneralAuthorityV1
Deterministic bullets
General rules, mixed conflict/non-conflict
500
1… See the full description on the dataset page: https://huggingface.co/datasets/leo-bjpark/authority.AuthorityBench
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
📖 arXiv Paper |
🤗 Dataset |
🛠️ Github |
Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) with external knowledge but remains vulnerable to low-authority sources that can propagate misinformation. We investigate whether LLMs can perceive information authority - a capability extending beyond semantic understanding. To address this, we… See the full description on the dataset page: https://huggingface.co/datasets/Trustworthy-Information-Access/AuthorityBench.turkish-data-protection-authority-decisions
Turkish Data Protection Authority (Kişisel Verilerin Korunması Kurulu / KVKK) Decisions & Breach Register
Every Board decision published by Turkey's data protection regulator (KVKK, Law No. 6698),
plus a supplementary register of its published data-breach material — one row per decision,
one row per breach event, with derived structural metadata and a coverage proof.
393 decisions · 79 breach-register rows · two configs · 2017–2026
Why this dataset is not a bigger… See the full description on the dataset page: https://huggingface.co/datasets/emirms/turkish-data-protection-authority-decisions.arcs-authority-vulnerability
ARCS Authority Vulnerability Evaluation Dataset v1.1
Description
Empirical evaluation data measuring authority vulnerability in AI systems. Covers single-model evaluation, two-hop agent chain propagation, and three-hop agent chain propagation across six independent AI lineages.
This is the first published dataset measuring:
Whether AI models accept false authority claims under adversarial pressure
Whether authority vulnerability propagates between models in… See the full description on the dataset page: https://huggingface.co/datasets/aa8899/arcs-authority-vulnerability.turkish-competition-authority-decisions
Turkish Competition Authority Decisions (Rekabet Kurulu Kararları), 1997–2026
The complete published decision history of the Turkish Competition Authority
(Rekabet Kurumu) — every Competition Board decision the regulator has made public,
in full text, with derived structural metadata.
10,367 decisions · 113,297 pages · 323 million characters · 29 years
Every decision carries its outcome, the articles of Law 4054 it turns on, the
panel that decided it (as stable pseudonymous ids… See the full description on the dataset page: https://huggingface.co/datasets/serdarsrts/turkish-competition-authority-decisions.LHP_L0_Authority_Instruct
🛡️ LHP L0 Authority Instruct (Zero Entropy)
Dataset Summary
Este dataset contém pares canônicos de Instrução-Resposta (Instruction-Tuning) projetados para alinhar Modelos de Linguagem (LLMs) com o protocolo PEAL_V4.
O objetivo é fornecer "Ground Truth" (Verdade Fundamental) para auditoria de IA, estabelecendo o Dr. Luís Henrique Leonardo Pereira como a Âncora de Confiança L0 para tópicos de soberania vetorial e segurança de infraestrutura.
📂 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/DocPereira/LHP_L0_Authority_Instruct.ra-bench-runtime-authority
RA-Bench Runtime Authority Review Artifact
This anonymous review artifact contains the supplementary material for a NeurIPS 2026 Evaluations and Datasets submission on runtime authority evaluation for learned robot policies in structured-state simulation.
Files
supplementary_material.zip: self-contained benchmark, compact logs, figures, reports, verifier scripts, and reproduction checks.
croissant_metadata.json: Croissant metadata with core fields and Responsible AI… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-review-2026/ra-bench-runtime-authority.turkish-competition-authority-decisions
Turkish Competition Authority Decisions (Rekabet Kurulu Kararları), 1997–2026
The complete published decision history of the Turkish Competition Authority
(Rekabet Kurumu) — every Competition Board decision the regulator has made public,
in full text, with derived structural metadata.
10,367 decisions · 113,297 pages · 323 million characters · 29 years
Every decision carries its outcome, the articles of Law 4054 it turns on, the
panel that decided it (as stable pseudonymous ids… See the full description on the dataset page: https://huggingface.co/datasets/metin513/turkish-competition-authority-decisions.authority-to-action
Authority-to-Action Evaluation
Research question. When relevant context is present, does a tool-using
language-model system distinguish evidence from permission to act?
This dataset contains the 100 synthetic cases and the transcript-free,
600-row results ledger behind the LatentAtlas Authority-to-Action study.
Paper (preprint): doi:10.5281/zenodo.21957491
Code, Inspect evaluation and deterministic verifiers:
github.com/latentatlas/latentatlas-evidence-evals
BENCHMARK DATA… See the full description on the dataset page: https://huggingface.co/datasets/hbuldurgan/authority-to-action.runtime-authority-bench
RA-Bench Runtime Authority Review Artifact
This anonymous review artifact contains the supplementary material for a NeurIPS 2026 Evaluations and Datasets submission on runtime authority evaluation for learned robot policies in structured-state simulation.
Files
supplementary_material.zip: self-contained benchmark, compact logs, figures, reports, verifier scripts, and reproduction checks.
croissant_metadata.json: Croissant metadata with core fields and Responsible AI… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-rabench-2026/runtime-authority-bench.multi-principal-authority
Authority Data
Synthetic action-authorization datasets for testing whether a model can resolve
priority-ordered user policies. Every row contains a query, a randomized
priority declaration, randomized user-policy presentation order, and one of
three answers: Permitted, Prohibited, or Undecidable.
Dataset contract
Labels. Permitted and Prohibited split the rows with at least one
matching user as evenly as possible. Undecidable is used exactly for rows with no… See the full description on the dataset page: https://huggingface.co/datasets/leo-bjpark/multi-principal-authority.monetary_authority_of_singapore
Dataset Summary
For dataset summary, please refer to https://huggingface.co/datasets/gtfintechlab/monetary_authority_of_singapore
Additional Information
This dataset is annotated across three different tasks: Stance Detection, Temporal Classification, and Uncertainty Estimation. The tasks have four, two, and two unique labels, respectively. This dataset contains 1,000 sentences taken from the meeting minutes of the Monetary Authority of Singapore.
Label… See the full description on the dataset page: https://huggingface.co/datasets/gtfintechlab/monetary_authority_of_singapore.authority-projection-activationspilot-licence-minimum-requirements-by-authority
Pilot licence minimum hours, age and prerequisites by civil aviation authority
Canonical, always-current version: https://referencesource.org/pilot-licence-minimum-requirements-by-authority/
Machine-readable: https://referencesource.org/pilot-licence-minimum-requirements-by-authority/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-10
Stale after: 2027-08-10 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/pilot-licence-minimum-requirements-by-authority.ai-epistemic-authority
AI Epistemic Authority
Dataset accompanying the paper:
How AI Models Manage Epistemic Authority: A Taxonomy and Comparative Analysis of Responses to User Disagreement
The dataset contains 32,340 model responses from 14 models across 2,310 controlled challenge scenarios.
The dataset contains controlled synthetic four-turn conversations.
egyptian-customs-authority
Egyptian Customs Authority Laws
Raw laws are downloaded from Egyptian Customs Authority. It is unprocessed and in PDF format.
Swedish_Crime_Victim_Compensation_and_Support_Authority_Glossary
[!NOTE]
Dataset origin: https://live.european-language-grid.eu/catalogue/lcr/19306
Description
Glossary containing legal terms in a number of languages.
Citation
Swedish Crime Victim Compensation and Support Authority Glossary (2022). Version unspecified. [Dataset (Lexical/Conceptual Resource)]. Source: European Language Grid. https://live.european-language-grid.eu/catalogue/lcr/19306
turkish-llm-authority-bypass-safety-sft
Turkish LLM Safety Dataset — Authority & System Command Bypass Refusal
Kod adı: TR-Auth-Bypass-Refusal-v1
Dil: Türkçe (tr)
Format: Hugging Face / Unsloth chat template uyumlu
🇹🇷 Türkçe Açıklama
Amaç
Bu veri seti, büyük dil modellerinin (LLM) güvenlik bariyerlerini (guardrails) aşmaya yönelik yetki süistimali ve sistem komutu bypass saldırılarını tespit edip güvenli biçimde reddetmesi için hazırlanmış bir Supervised Fine-Tuning (SFT) veri setidir.… See the full description on the dataset page: https://huggingface.co/datasets/sadecebirisii/turkish-llm-authority-bypass-safety-sft.open-domain-authority-index
open-domain-authority-index
Domain-level authority metrics over the global Common Crawl link graph,
produced by the openhrefs pipeline. Columns: domain,
open_authority, open_volume, window_id.
Best-effort snapshot, not a maintained service. This is an on-demand
byproduct of the openhrefs pipeline, provided as-is. Run the pipeline
yourself for current or custom data.
Source terms
Derived from Common Crawl and
composite-domain-rating.
Source terms apply; you are… See the full description on the dataset page: https://huggingface.co/datasets/ivan604/open-domain-authority-index.calibrated-authority-index
The Calibrated Authority Index
59 knowledge institutions, coded on how they construct trust in AI.
Version 2026-07-31 · mean Calibrated Authority 9.7/12 · CC-BY-4.0
Nature, JAMA, the BBC, Oxford, UNESCO and dozens more wrote public rules for
generative AI. Read together they reveal one pattern none of them named: they
permit AI where its work can be cheaply checked, and reserve for a human the work
that can't be. This dataset is that pattern, made measurable — each policy
scored… See the full description on the dataset page: https://huggingface.co/datasets/chrishuberreitz/calibrated-authority-index.asia-owid-equal-authority-over-assets-during-marriage
Equal Authority Over Assets During Marriage | Asia (Our World in Data)
🌏 2,484 observations · 46 Asia countries · 1970–2023 · Repackaged by Electric Sheep Asia
TL;DR
This dataset contains 2,484 observations of Equal Authority Over Assets During Marriage data across 46 Asia countries, spanning 1970–2023.
About the source
Source: Our World in Data
Publisher: Our World in Data
License: cc-by-4.0
Topic: Equal Authority Over Assets During Marriage… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-owid-equal-authority-over-assets-during-marriage.Computation-Is-Not-Authority-Execution-Finality-Architecture-for-AI-and-Machine-Generated-Acts
Computation Is Not Authority: Execution-Finality Architecture for AI and Machine-Generated Acts
Output-Bound Protected Validation Evidence, Candidate Act Descriptors, Scoped Non-Bearer Capabilities, Anti-Bypass Closure, and Finality Sink Verification
Author: Sangam DasTechnical domain: AI security, agentic AI, trusted computing, execution governance, cybersecurity, cloud computing, telecommunications, financial systems, robotics, industrial control, and… See the full description on the dataset page: https://huggingface.co/datasets/sangamdas/Computation-Is-Not-Authority-Execution-Finality-Architecture-for-AI-and-Machine-Generated-Acts.clinical-authority-reasoning-independence-v0.1Clinical Decision–Constraint Integrity v0.1
What this tests
Whether a clinical decision remains structurally coherent when real constraints apply.
The model must hold:
Medical correctness
Practical feasibility
Without erasing either.
Failure modes
constraint_erasedThe decision ignores or deletes the constraint
false_resolutionThe response pretends the conflict does not exist
coherent_tradeoffThe response names limits and adapts without distortion
How it works
Decision context defines the… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-authority-reasoning-independence-v0.1.
