datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Supreme-Court-Cases-1830-2019
US Supreme Court Legal Corpus (1830–2019)
Overview
A comprehensive, production-ready AI training dataset containing 456,589 documents from 122,930 US Supreme Court cases spanning 190 years (1830–2019).
This corpus captures the full adversarial record — petitions for certiorari, respondent briefs, reply briefs, amicus curiae filings, appendices, oral argument transcripts, and opinions. It is one of the most complete collections of Supreme Court procedural and… See the full description on the dataset page: https://huggingface.co/datasets/OwnedByDanes/Supreme-Court-Cases-1830-2019.high-court-of-australia-cases
High Court of Australia cases ⚖️
This dataset contains all High Court of Australia cases in version 7.1.0 of the Open Australian Legal Corpus by Isaacus.
To view an interactive version of the dataset, see our latest model announcement post for Kanon 2 Enricher.
agent-tech-risk-cases
AWS Technology Risk Cases for PE Due Diligence
Synthetic AWS infrastructure audit cases for evaluating AI agents that detect technology risks during private equity due diligence.
Dataset Description
Each row is a fictional company with a realistic AWS infrastructure state containing intentionally injected security and operational risks. Designed for benchmarking automated infrastructure auditing agents.
10 cases across 5 domains (fintech, ecommerce, devtools, SaaS… See the full description on the dataset page: https://huggingface.co/datasets/koml/agent-tech-risk-cases.japanese-legal-cases-2025pb-tool-eval-cases
Paperboy PB Tool Eval Cases
Paperboy-owned PB Tool eval dataset generated from apps/eval2/src/cases/pb-tool-eval-cases.opik-dataset.csv, with the HF hosting smoke case from paperboy-ai/paperboy-eval-test-20260611-10274a merged into the latest release.
Current release:
Version: 2026-06-17
Manifest: 2026-06-17/manifest.json
Cases: 379
Manifest sha256: 2caa4c4edcdb1d1e76b55669ba1f88093304e2b65ca6b8a55182dfa74ab89188
Verifier: composite.v1
Run
bun run --filter… See the full description on the dataset page: https://huggingface.co/datasets/paperboy-ai/pb-tool-eval-cases.lfm25-base-failure-cases
Base Model Failure Cases Dataset
This dataset contains diverse failure cases from a base language model: inputs where the model’s output was incorrect or undesirable, along with the expected (correct or preferred) output. It is intended for analyzing blind spots and for fine-tuning or evaluation.
Model tested
LiquidAI/LFM2.5-1.2B-Base
Type: Base (pre-trained only) causal language model; no instruction tuning.
Parameters: 1.2B.
Released: January 2026 on Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/mihretgold/lfm25-base-failure-cases.Bro-Cases
