datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
legal-scenarios-SCOTUS-2024-decisions
Purpose and scope
This dataset evaluates an LLM's reasoning ability in a legal context. Each question presents a realistic scenario involving competing legal principals,
and asks the LLM to present a correct legal resolution with sufficient justification based on precedent. The dataset was created using slip opinions of
the US Supreme Court from the 2024 term, taken from the Supreme Court website.
Dataset Creation Method
The benchmark was created using RELAI’s data… See the full description on the dataset page: https://huggingface.co/datasets/relai-ai/legal-scenarios-SCOTUS-2024-decisions.bva-decisions-structured-sample2019Present
BVA Structured Decisions (2019–2025)
Structured, issue-level records extracted from U.S. Board of Veterans' Appeals (BVA) decisions — each decision parsed into its issues, conditions, outcomes, citations, and reasoning, with per-document provenance and completeness flags. Built for training and evaluating legal-AI models on veterans' disability adjudication.
This is a 2900-decision sample, balanced across seven years (2019–2025, decisions/year), so it's representative of the… See the full description on the dataset page: https://huggingface.co/datasets/williamTLmiller/bva-decisions-structured-sample2019Present.decisiones-de-seguridad-pabloconfigs:
config_name: main_data
data_files: "decisiones2.csv"
default: true
