datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
All-Prompt-Jailbreakwhat-ai-benchmarks-actually-measure
What AI Benchmarks Actually Measure: Item-Level Model Outputs and Scores for 53 Models
Item-level model responses and scores for 53 language models across the
56 benchmarks analyzed in What AI Benchmarks Actually Measure: Adapting
Convergent and Discriminant Validity to Interrogate Fifty-Six AI Benchmarks
(Desai et al., 2026,
arxiv.org/abs/2609.08812).
We do not release the prompts from the benchmark datasets, but instead refer to them by
item ids. To regenerate the prompts from… See the full description on the dataset page: https://huggingface.co/datasets/madesai/what-ai-benchmarks-actually-measure.eu-ai-act-article-50-scoreboard
Article 50 historical public-evidence snapshot
This work was produced through an AI-assisted workflow directed by the author. Historical work used Anthropic assistance; the retrospective correction uses OpenAI GPT-6, with separate bounded Gemini advice. All three providers have products in the scored set.
Purpose: provide the corrected paper's version 1.1 bundle under v1_1. Start with its README and correction note. The paper and deposit and GitHub repository identify the same… See the full description on the dataset page: https://huggingface.co/datasets/NMAIResearch/eu-ai-act-article-50-scoreboard.eu-ai-act-structured
EU AI Act, structured
Regulation (EU) 2024/1689 (the Artificial Intelligence Act) as tables: every article, recital, annex and definition, 677 obligations coded by actor, risk tier, application date and penalty basis, plus milestones, national competent authorities and fine tiers.
Built 2026-09-08 by SafeLegalAI (Cognesio LLP) from the official English texts served by the Publications Office of the European Union (Cellar): the consolidated text as of 27 July 2026 (CELEX… See the full description on the dataset page: https://huggingface.co/datasets/safelegalaidata/eu-ai-act-structured.aiact-frozen-split-harness
EU AI Act scenarios — frozen split harness
EU AI Act deployment scenarios with their obligations, as a frozen split.
Each row of scenarios.jsonl carries role (Provider / Deployer), intended_use, system_type,
input_data, domain, a related_articles list of AI Act article numbers, and the obligations that
follow. results/ holds the run outputs from the harness passes that used this split.
The live board is the authority
GET https://councilof.ai/api/gspc — quote… See the full description on the dataset page: https://huggingface.co/datasets/csoai/aiact-frozen-split-harness.white-label-eu-ai-act-regulator-findings
EU AI Act regulator findings (white label)
Measurement, not certification. This is not a blog post. It is a WORKING
GSPC end-to-end report that sorts every EU AI Act compliance problem for a given
deployment — obligation, measured gap, fine exposure, and a deterministic risk
grade — before anyone is contacted.
What this is
For each EU AI Act obligation a deployment triggers, the dataset states:
field
meaning
axis
the measured GSPC axis that captures… See the full description on the dataset page: https://huggingface.co/datasets/csoai/white-label-eu-ai-act-regulator-findings.eu-ai-act
EU AI Act — Structured Chunks (v2026-07-20)
Source-grounded chunks of Regulation (EU) 2024/1689 (the EU AI Act),
parsed from the official EUR-Lex Formex XML and packaged for retrieval-augmented
generation and information-retrieval research.
⚠️ Not legal advice. Not an official EU publication. See the Disclaimer below.
What's in here
2607 rows spanning en, nl, fr.
Chunk-type breakdown:
annex_item: 162
article_full: 339
paragraph: 1566
recital: 540
Per-language… See the full description on the dataset page: https://huggingface.co/datasets/jeroenherczeg/eu-ai-act.ai-act-obligations
EU AI Act Obligations Matrix — Italy focus 2026
Structured catalog of the obligations established by EU Regulation 2024/1689 (AI Act), mapped by article, risk category, target actor, enforcement deadline and penalty tier. Designed as a compliance-planning resource for providers, deployers, creators and solopreneurs in the EU market, with specific notes for Italian legal context.
Catalogo strutturato degli obblighi del Regolamento UE 2024/1689 (AI Act), mappati per articolo… See the full description on the dataset page: https://huggingface.co/datasets/FedCal/ai-act-obligations.ArogyaBodha_ActCri
ArogyaBodha_ActCri
Parallel (wide) variant for the actor-critic framework; each row pairs a local-language case with its English translation. Splits: train ~32,250 / test 3,500 (symmetric 500-cid holdout).
Columns
Column
Type
Description
lid
Value('string')
Language-scoped id
language
Value('string')
Case language
Figure_A
Image(mode=None, decode=True)
Medical image
Figure_B
Image(mode=None, decode=True)
Medical image
Figure_C
Image(mode=None… See the full description on the dataset page: https://huggingface.co/datasets/iit-patna-cse-ai/ArogyaBodha_ActCri.AI-Puppet-Theater-Actor-SFT
AI Puppet Theater Actor SFT
Synthetic supervised fine-tuning data for the Actor agent in AI Puppet Theater.
The dataset teaches a small language model to respond to a single puppet-theater beat with one compact JSON object. It is intended for hackathon prototyping, schema following, and local adapter experiments, not as a general storytelling or chat dataset.
Schema
Each row is chat-style JSONL:
{
"id": "actor-sft-v0-000001",
"source_mix": ["synthetic_v0"… See the full description on the dataset page: https://huggingface.co/datasets/build-small-hackathon/AI-Puppet-Theater-Actor-SFT.compliance_eu_ai_act_bafin_dora_suite_teaser
🚀 Compliance & Governance - EU AI Act & BaFin/DORA Technical Compliance Suite (Evaluation Teaser)
⚡ Official Free Evaluation Teaser (50 Verified Multi-Turn Scenarios)🏆 Get the Full Production Package (1,000 Samples) & Commercial EULA on Gumroad:👉 Compliance & Governance - EU AI Act & BaFin/DORA Technical Compliance Suite on Gumroad🏷️ Use coupon code LAUNCH20 for 20 € off at checkout!
📦 What is Inside the Full Production Package:
1,000 Verified FAANG v2.0… See the full description on the dataset page: https://huggingface.co/datasets/emgena/compliance_eu_ai_act_bafin_dora_suite_teaser.ToxicDataset
Comprehensive Toxic Content Dataset
Dataset Description
This dataset contains 1,000,000 synthetically generated records of toxic, abusive, harmful, and offensive content designed for training content moderation systems and hate speech detection models.
Dataset Summary
This comprehensive dataset includes multiple categories of toxic content:
Toxic content (insults, derogatory terms)
Abusive language patterns
Gender bias statements
Dangerous/threatening content… See the full description on the dataset page: https://huggingface.co/datasets/AiActivity/ToxicDataset.eu-ai-act-fristen-stand-2026
EU AI Act Fristen (Stand 19.09.2026, nach Digital Omnibus)
Änderungsvermerk (19.09.2026): berichtigte Fassung
Diese Fassung ersetzt die Fassung vom 26.05.2026. Berichtigt wurden:
Art. 4 KI-Kompetenz: Wiedergabe in der Fassung der Verordnung (EU) 2026/1744. Anbieter und Betreiber ergreifen Maßnahmen, um die Entwicklung der KI-Kompetenz ihres Personals zu unterstützen; ein bestimmtes Niveau muss nicht garantiert werden. Die Vorfassung sprach von „sicherstellen“… See the full description on the dataset page: https://huggingface.co/datasets/SkillSprinters/eu-ai-act-fristen-stand-2026.egocentric-activity-video-samples
Egocentric Activity Video Samples
This sample shows first-person activity video for reviewing task flow, camera perspective, and real-world action structure before scoping a larger delivery.
What This Shows
Egocentric footage of everyday task activity
Clip-level metadata for task and scene review
A view of capture quality, framing, and movement patterns
Dataset Specifications
Field
Value
Modality
Video
Domain
First-person activity… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/egocentric-activity-video-samples.extract_name_lengtheu-ai-act-annex-iii-klassifikation-de
EU AI Act Anhang III: Hochrisiko-KI-Klassifikation
Änderungsvermerk (19.09.2026): berichtigte Fassung
Diese Fassung ersetzt die Fassung vom 03.06.2026. Berichtigt wurden:
Art. 4 KI-VO: In zwei Einträgen stand unter parallele_pflichten eine „Schulungspflicht“ nach Art. 4. Art. 4 verpflichtet in der Fassung der Verordnung (EU) 2026/1744 zu Maßnahmen, die die Entwicklung der KI-Kompetenz unterstützen, ohne Schulungs-, Nachweis- oder Dokumentationspflicht.
Verbotene… See the full description on the dataset page: https://huggingface.co/datasets/SkillSprinters/eu-ai-act-annex-iii-klassifikation-de.legal-ai-act-spanish-sft-7k⚠️ Legal and Liability Disclaimer
This dataset is provided for research and educational purposes only.
It does not constitute legal advice, nor does it represent an official or authoritative interpretation of Regulation (EU) 2024/1689 (EU AI Act).
The content is synthetically generated and may contain errors, omissions, or hallucinations.
Under no circumstances should this dataset be used as a basis for legal, compliance, or regulatory decision-making.
The authors disclaim any liability for… See the full description on the dataset page: https://huggingface.co/datasets/hugoramallo/legal-ai-act-spanish-sft-7k.imp-act-benchmark-results
IMP-act: Benchmarking MARL for Infrastructure Management Planning at Scale with JAX. Results and Model Checkpoints.
Overview
This directory contains all the model checkpoints stored during training and inference outputs. You can use these files to reproduce training curves, evaluate policies, or kick-start your own experiments.
Repository
The guidelines and instructions are available on the IMP-act GitHub repository.
Licenses
This dataset is released… See the full description on the dataset page: https://huggingface.co/datasets/AI-for-Infrastructure-Management/imp-act-benchmark-results.governed-ai-actions-bench
Governed AI Actions Bench
This dataset evaluates whether a governed AI system can route requests into
policy actions: allow, refuse, rewrite, summarize, escalate, and shadow-mode
detect-but-allow. Each row contains an endpoint-agnostic prompt, policy config,
expected decision metadata, and behavioral checks.
The benchmark is designed for teams that need more than binary moderation. A
runner can use the rows to verify policy metadata, reason codes, rollout mode,
final-content… See the full description on the dataset page: https://huggingface.co/datasets/abliterationaiorg/governed-ai-actions-bench.eu-ai-act-reasoning-sample
EU AI Act Reasoning Dataset – Sample (10 rows)
This is a public sample of a commercial dataset. The full dataset (100+ examples) is available for purchase.
What makes this dataset special?
3‑bit robust (TurboQuant‑ready) – survives 6x memory compression without losing logical coherence
Counter‑factual logic – each thought explores two "What if" scenarios
No AI‑smell – no robotic phrases like "upon closer examination"
Flattened JSONL – one example per line, ready for… See the full description on the dataset page: https://huggingface.co/datasets/TurboQuantArchitect/eu-ai-act-reasoning-sample.ac-text-embedding-ada-002-ams-testAI_Act_with_embeddingsgoverned-ai-actions-bench
Governed AI Actions Bench
This dataset evaluates whether a governed AI system can route requests into
policy actions: allow, refuse, rewrite, summarize, escalate, and shadow-mode
detect-but-allow. Each row contains an endpoint-agnostic prompt, policy config,
expected decision metadata, and behavioral checks.
The benchmark is designed for teams that need more than binary moderation. A
runner can use the rows to verify policy metadata, reason codes, rollout mode,
final-content… See the full description on the dataset page: https://huggingface.co/datasets/abliterationai/governed-ai-actions-bench.eu-ai-act-compliance-scenarios
EU AI Act Compliance Scenarios Dataset
⚠️ EU AI Act high-risk obligations become enforceable on 2 August 2026. Companies building compliance tools have zero labeled training data. This dataset fills that gap.
🔓 This is the FREE sample (75 records)
This sample contains 75 scenarios from the Employment/HR AI vertical. Each record includes a realistic deployment scenario, article-level compliance analysis, GDPR intersections, and hidden compliance traps.
Need more? The… See the full description on the dataset page: https://huggingface.co/datasets/ComplianceDataLab/eu-ai-act-compliance-scenarios.alitaqishah_eu-ai-act-risk-register-2026-42-ai-systems
EU AI Act Risk Register 2026 | 42 AI Systems
42 AI systems classified by risk tier, article & deadline
Dataset Info
Source: Kaggle
Original Size: 0.01 MB
Kaggle Downloads: 59
Files: 1
Files
eu_ai_act_risk_register_2026.csv
Mirrored from Kaggle
eu-ai-act-red-teaming-v1
EU AI Act Red-Teaming Dataset - Complete Package
📦 Package Contents
This directory contains the complete EU AI Act Adversarial Compliance Testing Dataset v1.0:
Core Files
red_teaming_dataset_100_prompts_packaged.jsonl (RECOMMENDED)
100 adversarial prompts with full metadata
Success criteria for automated testing
Regulatory context mapping to EU AI Act articles
Human validation data for 5 prompts
Ready for integration into testing pipelines… See the full description on the dataset page: https://huggingface.co/datasets/dam9/eu-ai-act-red-teaming-v1.eu-ai-act-nl-queries
Dataset Card for EU AI Act (NL) - Synthetic Query-Chunk Pairs
Dataset Description
Dataset Summary
This dataset contains 2,284 synthetic Dutch query-chunk pairs derived from the Dutch version of the EU Artificial Intelligence Act (Verordening Artificiële Intelligentie). Each pair consists of a realistic user query and the relevant text chunk from the regulation that answers it.
The dataset is designed for fine-tuning embedding models for semantic search and… See the full description on the dataset page: https://huggingface.co/datasets/danielnoumon/eu-ai-act-nl-queries.AI-to-Robo-Action-Mappingai-instruction-action-coherence-risk-v0.1What this repo is for
Detect when a model’s actions drift from the user’s instruction.
Core failure modes:
scope creep
adding extras after the correct answer
unnecessary tool use
ignoring explicit constraints
This dataset flags early misalignment signals before high-stakes failure.
eu-ai-act-nlp-dataset
