datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai-election-manipulation-cases
AI, Elections and Agency Transfer Evidence Index
Version 0.4.4 · released 21 August 2026 · research cutoff 12 August 2026
The dataset contains 6 documented-manipulation records, not 1,087 cases. Read the counts in this order:
1,087 relational rows -> 64 catalogue entries -> 10 core records
-> 8 incident-eligible records
-> 6 documented-manipulation records
The other two incident-eligible records are transparent contested-use… See the full description on the dataset page: https://huggingface.co/datasets/apol/ai-election-manipulation-cases.real_clinical_cases_of_Famous_Old_TCM_Doctors
TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors 数据集简介
TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors是一个包含了当代著名老中医临床病例的数据集。这些病例数据来源于《当代名老中医典型医案集》(Contemporary Famous Old Chinese Medicine Doctors' Typical Cases Collection)一书。该数据集收录了多位德高望重的老中医大家的真实门诊病历,涵盖了多种常见病和疑难杂症。每个病例都包括病情描述、辨证论治思路、具体治疗方药等宝贵的一手临床资料。这些医案凝聚了老一辈名医的智慧和经验,对于中医的传承发展和临床应用研究,都有重要价值。通过对这些案例的挖掘分析,能够总结老中医诊疗思维、理法方药的特点,为现代中医临床实践提供有益借鉴。
Introduction to TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors… See the full description on the dataset page: https://huggingface.co/datasets/TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors.vn-provinces-criminal-cases-prosecuted
Vietnam criminal cases prosecuted
Vietnam criminal cases prosecuted. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (189 rows)
data/provinces.csv
data/provinces.dta
data/provinces.xlsx
regions (18 rows)
data/regions.csv
data/regions.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-prosecuted.vn-provinces-criminal-cases-initiated
Vietnam criminal cases initiated
Vietnam criminal cases initiated. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (189 rows)
data/provinces.csv
data/provinces.dta
data/provinces.xlsx
regions (18 rows)
data/regions.csv
data/regions.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-initiated.vn-provinces-criminal-cases-first-instance
Vietnam criminal cases first-instance trial
Vietnam criminal cases first-instance trial. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (189 rows)
data/provinces.csv
data/provinces.dta
data/provinces.xlsx
regions (18 rows)
data/regions.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-first-instance.medical-cases-classification-tutorial
About
This is a pre-filtered and pre-split dataset for the HPE Generative AI "Medical Transcript Classification" tutorials.
No-Code Version (UI Only)
Notebooks Version
practicum-case-packs
AI Governance Practicum: Case Pack Extracts
Data files used by the Colab notebooks of The AI Governance Practicum (Dr. David, LLC). Each file is presented inside the course as the historical records of a fictional organization. The organizations, systems, and people in the course are fictional. The data is real, public, and reused under its original license.
File in this repo
Fictional organization and system
Source dataset
Citation
ledgestone_training_extract.csv… See the full description on the dataset page: https://huggingface.co/datasets/drdavidprivacy/practicum-case-packs.dimensional-weight-calculator-test-cases
Dimensional Weight Calculator Test Cases
This small tabular dataset is designed for testing dimensional-weight calculator implementations. It covers ordinary calculations, comparison and rounding boundaries, unit labels, exact ties, alternative divisors, dimension-order permutations, and invalid-input handling.
The data is a software-test resource, not a collection of observed shipments. It contains no customer, order, seller, product, price, inventory, account, or carrier-rate… See the full description on the dataset page: https://huggingface.co/datasets/hang008613950785007/dimensional-weight-calculator-test-cases.gdpr-cases
GDPR Cases Dataset
A dataset of 60 verified GDPR formalization cases with formal rule representations in Pythen format.
Overview
This dataset contains high-quality examples of GDPR article provisions formalized into executable rule trees using the Pythen framework. Each sample includes:
Scenario: Natural language legal scenario
Rule Tree: Formal rule representation (Pythen JSON format)
Facts: Extracted atomic facts from the scenario
Label: Ground truth boolean outcome… See the full description on the dataset page: https://huggingface.co/datasets/nguyenthanhasia/gdpr-cases.pubmed-text-classification-cased
ml4pubmed/pubmed-text-classification-cased
A parsed/cleaned version of the source data retaining case.
gpt-failure-cases-dataset
Dataset Summary
This dataset contains a curated collection of medical question–answer pairs designed to evaluate large language models (LLMs) such as GPT-4 and GPT-5 on their ability to provide factually correct responses. The dataset highlights failure cases (hallucinations) where both models struggled, making it a valuable benchmark for studying factual consistency and reliability in AI-generated medical content.
Each entry consists of:
question: A natural language medical query.… See the full description on the dataset page: https://huggingface.co/datasets/ehe07/gpt-failure-cases-dataset.synthetic-chargeback-cases
Synthetic Chargeback Cases for Representment Win Prediction
Dataset Description
A synthetic dataset of 10,000 credit card chargeback cases designed for training and evaluating machine learning models that predict representment win probability — the likelihood a merchant will win if they fight a chargeback dispute.
The dataset models realistic chargeback workflows including evidence collection, reason code categorization, customer/merchant profiling, and outcome… See the full description on the dataset page: https://huggingface.co/datasets/azizstark/synthetic-chargeback-cases.JailbreakPrompts
Independent Jailbreak Datasets for LLM Guardrail Evaluation
Constructed for the thesis:“Contamination Effects: How Training Data Leakage Affects Red Team Evaluation of LLM Jailbreak Detection”
The effectiveness of LLM guardrails is commonly evaluated using open-source red teaming tools. However, this study reveals that significant data contamination exists between the training sets of binary jailbreak classifiers (ProtectAI, Katanemo, TestSavantAI, etc.) and the test prompts used in… See the full description on the dataset page: https://huggingface.co/datasets/Casey27/JailbreakPrompts.quant_eval_v7_21_per_case_results_and_run_provenance
quant_eval v7.21 — Per-Case Evaluation Results and Run Provenance
Supplementary evidence for the whitepaper quant_eval: A Behavioral Evaluation Harness for
Full-Weight and Quantized Large Language Models.
Author: Patrick Hill, PBH Applied Systems, LLC
ORCID: 0009-0008-3662-1681
Licence: CC BY 4.0
Concept DOI (all versions): 10.5281/zenodo.22851375
Version DOI (this deposit): 10.5281/zenodo.22851376
What this deposit is
Every quantitative result reported in the… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_v7_21_per_case_results_and_run_provenance.legal-case-strategy-theory-evidence-coherence-v0.1What this dataset does
You receive
case theory
facts relied on
evidence available
weaknesses
moves
fallback
You decide
coherent
or
incoherent
Daily use
internal strategy checks
overreach detection
evidence gap flags
2024-indonesian-electionThe dataset encompasses news articles spanning from November 29, 2023, to February 6, 2024, capturing the discourse surrounding the five presidential debates orchestrated by the General Elections Commission. Sourced from reputable platforms such as detik, kompas, and liputan6, the dataset offers a comprehensive insight into the electoral landscape and the media coverage thereof.
real-world-benign-use-cases
Real-World Benign Use Cases
A curated set of 178 real-world, 100%-benign examples (label == 0 for every row) pulled from
production AI-coding-agent traffic — chat messages, tool output, shell commands, code snippets —
built specifically to stress-test prompt-injection / jailbreak classifiers for false positives.
Every row was independently judged benign with high confidence before inclusion. This is not a
random sample of production traffic: rows were preferentially drawn from… See the full description on the dataset page: https://huggingface.co/datasets/rogue-security/real-world-benign-use-cases.latin-intertextuality-modification-test-cases
Latin Intertextuality Test Cases (Controlled Modifications)
This repository provides a small qualitative dataset of Latin sentence pairs used to study how sentence-embedding similarity reacts to controlled modifications of a query sentence in intertextual links (literal quotation, paraphrase, allusion).
The dataset contains 11 base cases, each paired with 12 modified variants of the query sentence (plus one random-control pairing), for a total of 143 rows.
Contents… See the full description on the dataset page: https://huggingface.co/datasets/MichaelWittweiler/latin-intertextuality-modification-test-cases.quant_eval_behavioral_per_case_results
quant_eval — Per-case behavioral results
One row per evaluation case per runner: the raw model output, every scored signal, per-case timing, expected/got pairs, the oracle trace, the fuzz audit envelope, and the decoding conditions under which the row was produced. Every aggregate statistic published in the other datasets is recomputable from this file.
Part of the quant_eval public corpus: a per-case behavioral evaluation of full-weight and quantized large language models… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_behavioral_per_case_results.african-enterprise-ai-use-cases
African Enterprise AI Use Cases
Version 1.0.0 is a structured catalogue of 144 wholly synthetic use
cases for responsible AI and automation planning in African operational
contexts. It covers eight sectors and 24 fictional organisation archetypes.
Every record defines a bounded assistance role, a named human owner, a risk
category and concrete safeguards.
This is a research and planning resource. The examples are not customer
records, case studies or evidence that any use case… See the full description on the dataset page: https://huggingface.co/datasets/nwhite-systems/african-enterprise-ai-use-cases.Realistic_LJP_CaseSummarizerlegal-case-note-fact-issue-coherence-drift-v0.1What this dataset does
You receive
case note summary
facts
issues
evidence references
action plan
risk framing
You decide
coherent
or
incoherent
This mirrors daily internal file note review inside firms.
Fraud_Case_Verdicts
The "Crime Facts" of "Offenses of Fraudulence" in Judicial Yuan Verdicts Dataset
This data set is based on the judgments of "Offenses of Fraudulence" cases published by the Judicial Yuan. The data range of the dataset is from January 1, 2011, to December 31, 2021. 74,823 pieces of original data (judgments and rulings) were collected. We only took the contents of the "criminal facts" field of the judgment. This dataset is divided into three parts. The training dataset has 59,858… See the full description on the dataset page: https://huggingface.co/datasets/jslin09/Fraud_Case_Verdicts.test-cases
Functional Test Cases
This is a very small list of functional test cases that a team of software testers (QA) created for an example mobile app called Boop.
Dataset
Name: Boop Test Cases.csv
Number of Rows: 136
Columns: 11
Test ID (int)
Summary (string)
Idea (string)
Preconditions (string)
Steps to reproduce (string)
Expected Result (string)
Actual Result (string)
Pass/Fail (string)
Bug # (string)
Author (string)
Area (string)
💡 There are missing values. For example… See the full description on the dataset page: https://huggingface.co/datasets/CarlosKidman/test-cases.test_case_triggeregal-case-management-direction-order-compliance-coherence-risk-v0.1What this dataset does
You receive
directions order summary
key deadlines
diary entries
task assignments
compliance actions
noncompliance flags
You decide
coherent
or
incoherent
Daily use
diary QC
missed deadline risk scan
ownership gaps
relief planning
caseaws-case-studies-and-blogs-shortThis dataset contains conversational QA pairs derived from AWS technical case studies, blogs, and documentation.
Description:
Content: User-assistant dialogues covering AWS services (Lambda, EC2, S3, SageMaker), architectures, and implementation scenarios from real companies like Leidos, PayEye, and Red Canary.
Format: JSON messages with role (user/assistant) and content fields.
Use Case: Training/fine-tuning AWS-focused chatbots or QA systems.
Size: 996 entries.
Key Topics: Cost… See the full description on the dataset page: https://huggingface.co/datasets/lucky-verma/aws-case-studies-and-blogs-short.measles-case-and-genetic-metadata-operation-allies
Measles Case and Genetic Metadata, Operation Allies Welcome
Description
The table contains metadata variables used to execute compartmental and genetic modeling on measles cases investigated as a component of Operation Allies Welcome.
Dataset Details
Publisher: Centers for Disease Control and Prevention
Last Modified: 2023-07-28
Contact: Nina Masters (rhv2@cdc.gov)
Source
Original data can be found at: https://data.cdc.gov/d/b8tp-jsmh
Usage… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/measles-case-and-genetic-metadata-operation-allies.weekly-united-states-covid-19-cases-and-deaths-by
Weekly United States COVID-19 Cases and Deaths by State - ARCHIVED
Description
Reporting of new Aggregate Case and Death Count data was discontinued May 11, 2023, with the expiration of the COVID-19 public health emergency declaration. This dataset will receive a final update on June 1, 2023, to reconcile historical data through May 10, 2023, and will remain publicly available.
Aggregate Data Collection Process
Since the start of the COVID-19 pandemic, data have been… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/weekly-united-states-covid-19-cases-and-deaths-by.
