CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OwnedByDanes /Supreme-Court-Cases-1830-2019 US Supreme Court Legal Corpus (1830–2019) Overview A comprehensive, production-ready AI training dataset containing 456,589 documents from 122,930 US Supreme Court cases spanning 190 years (1830–2019). This corpus captures the full adversarial record — petitions for certiorari, respondent briefs, reply briefs, amicus curiae filings, appendices, oral argument transcripts, and opinions. It is one of the most complete collections of Supreme Court procedural and… See the full description on the dataset page: https://huggingface.co/datasets/OwnedByDanes/Supreme-Court-Cases-1830-2019.tabulartext-generation10K<n<100K0 likes131 downloads5mo agoHugging Face02isaacus /high-court-of-australia-cases High Court of Australia cases ‍⚖️ This dataset contains all High Court of Australia cases in version 7.1.0 of the Open Australian Legal Corpus by Isaacus. To view an interactive version of the dataset, see our latest model announcement post for Kanon 2 Enricher. texttext-generation1K<n<10K4 likes89 downloads7mo agoHugging Face03koml /agent-tech-risk-cases AWS Technology Risk Cases for PE Due Diligence Synthetic AWS infrastructure audit cases for evaluating AI agents that detect technology risks during private equity due diligence. Dataset Description Each row is a fictional company with a realistic AWS infrastructure state containing intentionally injected security and operational risks. Designed for benchmarking automated infrastructure auditing agents. 10 cases across 5 domains (fintech, ecommerce, devtools, SaaS… See the full description on the dataset page: https://huggingface.co/datasets/koml/agent-tech-risk-cases.texttext-generationn<1K0 likes17 downloads8mo agoHugging Face04nguyenthanhasia /japanese-legal-cases-2025gatedtabulartext-classification10K<n<100K1 likes14 downloads1y agoHugging Face05paperboy-ai /pb-tool-eval-cases Paperboy PB Tool Eval Cases Paperboy-owned PB Tool eval dataset generated from apps/eval2/src/cases/pb-tool-eval-cases.opik-dataset.csv, with the HF hosting smoke case from paperboy-ai/paperboy-eval-test-20260611-10274a merged into the latest release. Current release: Version: 2026-06-17 Manifest: 2026-06-17/manifest.json Cases: 379 Manifest sha256: 2caa4c4edcdb1d1e76b55669ba1f88093304e2b65ca6b8a55182dfa74ab89188 Verifier: composite.v1 Run bun run --filter… See the full description on the dataset page: https://huggingface.co/datasets/paperboy-ai/pb-tool-eval-cases.texttext-generationn<1K0 likes8 downloads3mo agoHugging Face06mihretgold /lfm25-base-failure-cases Base Model Failure Cases Dataset This dataset contains diverse failure cases from a base language model: inputs where the model’s output was incorrect or undesirable, along with the expected (correct or preferred) output. It is intended for analyzing blind spots and for fine-tuning or evaluation. Model tested LiquidAI/LFM2.5-1.2B-Base Type: Base (pre-trained only) causal language model; no instruction tuning. Parameters: 1.2B. Released: January 2026 on Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/mihretgold/lfm25-base-failure-cases.texttext-generationn<1K0 likes6 downloads7mo agoHugging Face07TommyDIL /Bro-Casestexttext-generation1K<n<10K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.