CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01anivcsh /supreme-court-datadocument10K<n<100K1 likes2.9k downloads5mo agoHugging Face02vihaannnn /Indian-Supreme-Court-Judgements-Chunked Indian Supreme Court Judgements Chunked Executive Summary The dataset aims to address the chronic backlog in the Indian judiciary system, particularly in the Supreme Court, by creating a dataset optimized for legal language models (LLMs). The dataset will consist of pre-processed, chunked, and embedded textual data derived from the Supreme Court's judgment PDFs. Problem and Importance - Motivation Indian courts are overwhelmed with pending cases, with the… See the full description on the dataset page: https://huggingface.co/datasets/vihaannnn/Indian-Supreme-Court-Judgements-Chunked.textfeature-extraction10K<n<100K6 likes2.5k downloads2y agoHugging Face03wildphoton /courtlistener_opinionstext100K<n<1M1 likes2.3k downloads2y agoHugging Face04mrfg /turkish-court-decisions Türk İçtihat Korpusu — 11.045.085 Mahkeme Kararı Türkiye'nin kamuya açık mahkeme kararlarından derlenmiş, bilinen en büyük Türkçe hukuk metni veri seti. 11.045.085 karar, 31.5 milyar karakter düz metin (5.50 GB Parquet), 1962'den 2026'ya. Yargıtay, Danıştay, Anayasa Mahkemesi ve UYAP Emsal üzerinden yerel/istinaf mahkemeleri. Kapsam Kaynak Karar sayısı Yıl aralığı Metin Dosya Yargıtay (yargitay) 9.820.145 1997–2026 19.5 milyar karakter 17 Danıştay… See the full description on the dataset page: https://huggingface.co/datasets/mrfg/turkish-court-decisions.tabulartext-generation10M<n<100M5 likes1.3k downloads1mo agoHugging Face05kardosdrur /norwegian-courts Norwegian Courts Parallel corpus of Nynorsk and Bokmål from Norwegian Court transcriptions. The data originates from the OPUS project. textsentence-similarity1K<n<10K1 likes1.2k downloads3y agoHugging Face06vGassen /Dutch-Rechtspraak-court-casestext100K<n<1M1 likes958 downloads1y agoHugging Face07openlegaldata /court-decisions-germanygated Open Legal Data: Court Decisions Germany This dataset is a preprocessed version of an Open Legal Data data dump, spefically it contains German court decisions. The dataset was automatically generated and uploaded to the HF hub using oldp-toolkit. Available dumps Date Configs 2026-05-20 dump-20260520, dump-20260520-10k, dump-20260520-1k 2022-10-18 dump-20221018, dump-20221018-10k, dump-20221018-1k Data format Each dataset sample has the… See the full description on the dataset page: https://huggingface.co/datasets/openlegaldata/court-decisions-germany.texttext-generation100K<n<1M16 likes944 downloads4mo agoHugging Face08zalizedata /us-court-opinions-dockets-judges-dataset US Court Opinions Metadata, Dockets & Judges (CourtListener) 10M opinion clusters, 70M dockets and 16K judges from official CourtListener / Free Law Project bulk data as metadata + derived-signals tables — citation graph, company litigation profiles; no opinion full text. Part of the DataForge Open Data program — full production packages, free for academic and personal use. Canonical dataset page: https://data.zalize.com/datasets/us-court-opinions-dockets-judges-dataset… See the full description on the dataset page: https://huggingface.co/datasets/zalizedata/us-court-opinions-dockets-judges-dataset.tabulartext-classification10M<n<100M0 likes865 downloads1mo agoHugging Face09lvanews /ukrainian-court-decisions Ukrainian Court Decisions — Judgment Prediction A dataset of Ukrainian court decisions for case outcome prediction, extracted from the State Court Decisions Registry (ЄДРСР). Task Given the facts section (ВСТАНОВИВ) of a court decision, predict the judgment outcome: Label Ukrainian Description approved Задоволено Claim fully satisfied dismissed Відмовлено Claim dismissed partial Частково задоволено Claim partially satisfied Data… See the full description on the dataset page: https://huggingface.co/datasets/lvanews/ukrainian-court-decisions.texttext-classification100K<n<1M0 likes783 downloads2mo agoHugging Face10santoshtyss /us-court-cases Dataset Card for "us-court-cases" More Information needed text1M<n<10M5 likes708 downloads3y agoHugging Face11dougalldeepmind /2026-08-14-courtroom synth courtroom run — per-stage snapshots (resumable generation cache) field value experiment synth courtroom run — per-stage snapshots (resumable generation cache) date_generated 20260815_201700 constitution constitutions/claude_distilled_09_principles_mid_20260804/constitution.md source_repo https://github.com/Matthew-Bozoukov/Lessons_from_constituitional_AFT.git @ b992089ffec3dbc23ba676ef5c1ebad319937daa models per-stage models — see manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-14-courtroom.tabular10K<n<100K0 likes597 downloads25d agoHugging Face12labofsahil /Indian-Supreme-Court-Judgmentsdocument10K<n<100K1 likes535 downloads8mo agoHugging Face13Charlie019 /CourtSI-BenchStepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports Abstract Sports have long attracted broad attention as they push the limits of human physical and cognitive capabilities. Amid growing interest in spatial intelligence for vision-language models (VLMs), sports provide a natural testbed for understanding high-intensity human motion and dynamic object interactions. To this end, we present CourtSI, the first large-scale spatial… See the full description on the dataset page: https://huggingface.co/datasets/Charlie019/CourtSI-Bench.imagevisual-question-answering1K<n<10K2 likes455 downloads7mo agoHugging Face14ismatsamadov /azerbaijan-court-data Azerbaijan Court System Dataset The most comprehensive open dataset of Azerbaijan's judicial system — 1.64 million structured records and 1.54 million court decision PDFs (~160 GB) covering court decisions, active cases, scheduled hearings, court registries, judges, lawyers, and mediator organizations. Built for AI engineers, legal tech startups, and researchers who need real-world legal data at scale. Quick Start Load with Hugging Face datasets from datasets… See the full description on the dataset page: https://huggingface.co/datasets/ismatsamadov/azerbaijan-court-data.imagetext-classification1M<n<10M2 likes391 downloads6mo agoHugging Face15andrew-mitchel /tax-court-opinions Tax Court Opinions Text of United States Tax Court opinions, primarily from 1995 through September 2026, with a small number of earlier opinions back to 1986. The Tax Court publishes its opinions as PDF files; these were converted to text using pdfminer. Dataset Structure 14,848 rows, one per opinion. Columns: Column Type Description filename string Original filename, encoding year/month/day/type/name/pages/docket/judge year string Filing year (four… See the full description on the dataset page: https://huggingface.co/datasets/andrew-mitchel/tax-court-opinions.texttext-generation10K<n<100K1 likes363 downloads23d agoHugging Face16overthelex /indian-court-decisions Indian Court Decisions A large-scale dataset of Indian court decisions with full text, metadata, and outcome labels covering the Supreme Court of India and 25 High Courts (1950–2026). Dataset Summary Config Train Validation Test Total high_courts 11,682,776 1,459,319 1,457,934 14,600,029 supreme_court 40,044 4,990 5,019 50,053 Total 14,650,082 This is one of the largest publicly available legal NLP datasets, containing over 14.6 million… See the full description on the dataset page: https://huggingface.co/datasets/overthelex/indian-court-decisions.tabulartext-classification10M<n<100M1 likes356 downloads4mo agoHugging Face17ursishant /nyayashastra-court-judgments-corpus NyayaShastra Indian Court Judgments Corpus (12.4M Judgments) A comprehensive, curated dataset of 12.4 Million Indian Supreme Court, High Court, and District Court judgments. Optimized for high-speed columnar retrieval via Apache Parquet and DuckDB. Total Partitions: 248 Format: Apache Parquet (Snappy compressed) Columns: case_id, cnr, case_title, court_name_normalized, court_level, decision_date, cleaned_text, text_length texttext-retrieval10M<n<100M0 likes353 downloads1mo agoHugging Face18rusheeliyer /german-courts Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/rusheeliyer/german-courts.text1K<n<10K1 likes348 downloads3y agoHugging Face19Alptekinege /turkish-court-decisions Türk İçtihat Korpusu — 11.045.085 Mahkeme Kararı Türkiye'nin kamuya açık mahkeme kararlarından derlenmiş, bilinen en büyük Türkçe hukuk metni veri seti. 11.045.085 karar, 31.5 milyar karakter düz metin (5.50 GB Parquet), 1962'den 2026'ya. Yargıtay, Danıştay, Anayasa Mahkemesi ve UYAP Emsal üzerinden yerel/istinaf mahkemeleri. Kapsam Kaynak Karar sayısı Yıl aralığı Metin Dosya Yargıtay (yargitay) 9.820.145 1997–2026 19.5 milyar karakter 17 Danıştay… See the full description on the dataset page: https://huggingface.co/datasets/Alptekinege/turkish-court-decisions.tabulartext-generation10M<n<100M3 likes326 downloads23d agoHugging Face20Alexey5676 /russian-supreme-court-plenum-acts Plenum Resolutions of the Supreme Court of Russia (1961–2026) Every act published in the «Постановления Пленума» section of the Russian Supreme Court's website: 1,504 records — 1,503 plenum resolutions plus 1 meeting agenda — with full texts, metadata and the court's original attachments. Coverage 1961–2026; completeness verified against the court's own index at collection time (the section reported exactly 1,504 documents). Постановления Пленума ВС РФ — руководящие разъяснения… See the full description on the dataset page: https://huggingface.co/datasets/Alexey5676/russian-supreme-court-plenum-acts.documentsummarization1K<n<10K2 likes314 downloads9d agoHugging Face21overthelex /ukrainian-court-decisions Ukrainian Court Decisions — Judgment Prediction A dataset of Ukrainian court decisions for case outcome prediction, extracted from the State Court Decisions Registry (ЄДРСР). Task Given the facts section (ВСТАНОВИВ) of a court decision, predict the judgment outcome: Label Ukrainian Description approved Задоволено Claim fully satisfied dismissed Відмовлено Claim dismissed partial Частково задоволено Claim partially satisfied Data… See the full description on the dataset page: https://huggingface.co/datasets/overthelex/ukrainian-court-decisions.texttext-classification100K<n<1M0 likes308 downloads4mo agoHugging Face22JuDDGES /pl-court-raw Dataset Card for JuDDGES/pl-court-raw Dataset Summary The dataset consists of Polish Court judgments available at https://orzeczenia.ms.gov.pl/, containing full content of the judgments along with metadata sourced from official API and extracted from the judgment contents. This dataset contains raw data. For instruction dataset see JuDDGES/pl-court-instruct. For graph dataset see JuDDGES/pl-court-graph. Supported Tasks and Leaderboards The dataset can be… See the full description on the dataset page: https://huggingface.co/datasets/JuDDGES/pl-court-raw.tabular100K<n<1M0 likes295 downloads1y agoHugging Face23overthelex /pl-court-decisions Polish Court Decisions The largest open dataset of Polish court decisions: 2,830,029 decisions with full texts across all court levels. What Makes This Dataset Unique Source This dataset Best on HF (JuDDGES) Difference Common courts 437,446 437,450 (pl-court-raw) same source Administrative courts 1,899,852 ~1,800,000 (pl-nsa) same source Supreme Court + Constitutional Tribunal + KIO 492,731 0 +493K unique Total 2,830,029 ~2,237,450 +593K (+26%)… See the full description on the dataset page: https://huggingface.co/datasets/overthelex/pl-court-decisions.texttext-generation1M<n<10M2 likes271 downloads4mo agoHugging Face24PiotrSty /saos-polish-court-judgments SAOS Speeches Corpus — Orzeczenia Sądów Polskich Korpus orzeczeń sądowych z Systemu Analizy Orzeczeń Sądowych (SAOS), wygenerowany z oficjalnego API SAOS (www.saos.org.pl/api/dump/judgments). Statystyki Metryka Wartość Sądy sądy powszechne, Sąd Najwyższy, Naczelny Sąd Administracyjny, Trybunał Konstytucyjny Typy orzeczeń wyroki, postanowienia, uzasadnienia, zarządzenia, uchwały Rekordy 296,692 Znaki 5,895,441,856 Słowa 855,425,796 Tokeny… See the full description on the dataset page: https://huggingface.co/datasets/PiotrSty/saos-polish-court-judgments.tabular100K<n<1M0 likes251 downloads3mo agoHugging Face25Gyrevortex /turkish-court-decisions Türk İçtihat Korpusu — 11.045.085 Mahkeme Kararı Türkiye'nin kamuya açık mahkeme kararlarından derlenmiş, bilinen en büyük Türkçe hukuk metni veri seti. 11.045.085 karar, 31.5 milyar karakter düz metin (5.50 GB Parquet), 1962'den 2026'ya. Yargıtay, Danıştay, Anayasa Mahkemesi ve UYAP Emsal üzerinden yerel/istinaf mahkemeleri. Kapsam Kaynak Karar sayısı Yıl aralığı Metin Dosya Yargıtay (yargitay) 9.820.145 1997–2026 19.5 milyar karakter 17 Danıştay… See the full description on the dataset page: https://huggingface.co/datasets/Gyrevortex/turkish-court-decisions.tabulartext-generation10M<n<100M1 likes251 downloads1mo agoHugging Face26vihaannnn /Chunked-Indian-Supreme-Court-Judgements Indian Supreme Court Judgements Chunked texttoken-classification10K<n<100K1 likes245 downloads2y agoHugging Face27joelniklaus /brazilian_court_decisions Dataset Card for predicting-brazilian-court-decisions Dataset Summary The dataset is a collection of 4043 Ementa (summary) court decisions and their metadata from the Tribunal de Justiça de Alagoas (TJAL, the State Supreme Court of Alagoas (Brazil). The court decisions are labeled according to 7 categories and whether the decisions were unanimous on the part of the judges or not. The dataset supports the task of Legal Judgment Prediction. Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/joelniklaus/brazilian_court_decisions.texttext-classification1K<n<10K23 likes242 downloads4y agoHugging Face28WhissleAI /supreme-court-india-meta-speechaudio10K<n<100K0 likes233 downloads11mo agoHugging Face29serdarsrts /turkish-court-decisions-duplicate Türk İçtihat Korpusu — 11.045.085 Mahkeme Kararı Türkiye'nin kamuya açık mahkeme kararlarından derlenmiş, bilinen en büyük Türkçe hukuk metni veri seti. 11.045.085 karar, 31.5 milyar karakter düz metin (5.50 GB Parquet), 1962'den 2026'ya. Yargıtay, Danıştay, Anayasa Mahkemesi ve UYAP Emsal üzerinden yerel/istinaf mahkemeleri. Kapsam Kaynak Karar sayısı Yıl aralığı Metin Dosya Yargıtay (yargitay) 9.820.145 1997–2026 19.5 milyar karakter 17 Danıştay… See the full description on the dataset page: https://huggingface.co/datasets/serdarsrts/turkish-court-decisions-duplicate.tabulartext-generation10M<n<100M1 likes203 downloads27d agoHugging Face30rtarun789 /indian-court-decisions Indian Court Decisions A large-scale dataset of Indian court decisions with full text, metadata, and outcome labels covering the Supreme Court of India and 25 High Courts (1950–2026). Dataset Summary Config Train Validation Test Total high_courts 11,682,776 1,459,319 1,457,934 14,600,029 supreme_court 40,044 4,990 5,019 50,053 Total 14,650,082 This is one of the largest publicly available legal NLP datasets, containing over 14.6 million… See the full description on the dataset page: https://huggingface.co/datasets/rtarun789/indian-court-decisions.tabulartext-classification10M<n<100M0 likes196 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.