CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01apol /ai-election-manipulation-cases AI, Elections and Agency Transfer Evidence Index Version 0.4.4 · released 21 August 2026 · research cutoff 12 August 2026 The dataset contains 6 documented-manipulation records, not 1,087 cases. Read the counts in this order: 1,087 relational rows -> 64 catalogue entries -> 10 core records -> 8 incident-eligible records -> 6 documented-manipulation records The other two incident-eligible records are transparent contested-use… See the full description on the dataset page: https://huggingface.co/datasets/apol/ai-election-manipulation-cases.text1K<n<10K0 likes769 downloads1mo agoHugging Face02TCMLM /real_clinical_cases_of_Famous_Old_TCM_Doctors TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors 数据集简介 TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors是一个包含了当代著名老中医临床病例的数据集。这些病例数据来源于《当代名老中医典型医案集》(Contemporary Famous Old Chinese Medicine Doctors' Typical Cases Collection)一书。该数据集收录了多位德高望重的老中医大家的真实门诊病历,涵盖了多种常见病和疑难杂症。每个病例都包括病情描述、辨证论治思路、具体治疗方药等宝贵的一手临床资料。这些医案凝聚了老一辈名医的智慧和经验,对于中医的传承发展和临床应用研究,都有重要价值。通过对这些案例的挖掘分析,能够总结老中医诊疗思维、理法方药的特点,为现代中医临床实践提供有益借鉴。 Introduction to TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors… See the full description on the dataset page: https://huggingface.co/datasets/TCMLM/real_clinical_cases_of_Famous_Old_TCM_Doctors.tabularn<1K2 likes130 downloads3mo agoHugging Face03letrinhan /vn-provinces-criminal-cases-prosecuted Vietnam criminal cases prosecuted Vietnam criminal cases prosecuted. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Comparison Color key Files provinces (189 rows) data/provinces.csv data/provinces.dta data/provinces.xlsx regions (18 rows) data/regions.csv data/regions.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-prosecuted.tabularn<1K0 likes98 downloads5d agoHugging Face04letrinhan /vn-provinces-criminal-cases-initiated Vietnam criminal cases initiated Vietnam criminal cases initiated. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Comparison Color key Files provinces (189 rows) data/provinces.csv data/provinces.dta data/provinces.xlsx regions (18 rows) data/regions.csv data/regions.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-initiated.tabularn<1K0 likes85 downloads5d agoHugging Face05letrinhan /vn-provinces-criminal-cases-first-instance Vietnam criminal cases first-instance trial Vietnam criminal cases first-instance trial. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Comparison Color key Files provinces (189 rows) data/provinces.csv data/provinces.dta data/provinces.xlsx regions (18 rows) data/regions.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-first-instance.tabularn<1K0 likes80 downloads5d agoHugging Face06hpe-ai /medical-cases-classification-tutorial About This is a pre-filtered and pre-split dataset for the HPE Generative AI "Medical Transcript Classification" tutorials. No-Code Version (UI Only) Notebooks Version text1K<n<10K6 likes79 downloads3y agoHugging Face07drdavidprivacy /practicum-case-packs AI Governance Practicum: Case Pack Extracts Data files used by the Colab notebooks of The AI Governance Practicum (Dr. David, LLC). Each file is presented inside the course as the historical records of a fictional organization. The organizations, systems, and people in the course are fictional. The data is real, public, and reused under its original license. File in this repo Fictional organization and system Source dataset Citation ledgestone_training_extract.csv… See the full description on the dataset page: https://huggingface.co/datasets/drdavidprivacy/practicum-case-packs.tabular1K<n<10K1 likes70 downloads12h agoHugging Face08hang008613950785007 /dimensional-weight-calculator-test-cases Dimensional Weight Calculator Test Cases This small tabular dataset is designed for testing dimensional-weight calculator implementations. It covers ordinary calculations, comparison and rounding boundaries, unit labels, exact ties, alternative divisors, dimension-order permutations, and invalid-input handling. The data is a software-test resource, not a collection of observed shipments. It contains no customer, order, seller, product, price, inventory, account, or carrier-rate… See the full description on the dataset page: https://huggingface.co/datasets/hang008613950785007/dimensional-weight-calculator-test-cases.tabularn<1K0 likes54 downloads15d agoHugging Face09nguyenthanhasia /gdpr-cases GDPR Cases Dataset A dataset of 60 verified GDPR formalization cases with formal rule representations in Pythen format. Overview This dataset contains high-quality examples of GDPR article provisions formalized into executable rule trees using the Pythen framework. Each sample includes: Scenario: Natural language legal scenario Rule Tree: Formal rule representation (Pythen JSON format) Facts: Extracted atomic facts from the scenario Label: Ground truth boolean outcome… See the full description on the dataset page: https://huggingface.co/datasets/nguyenthanhasia/gdpr-cases.tabularn<1K0 likes47 downloads5mo agoHugging Face10ml4pubmed /pubmed-text-classification-cased ml4pubmed/pubmed-text-classification-cased A parsed/cleaned version of the source data retaining case. texttext-classification1M<n<10M0 likes45 downloads4y agoHugging Face11ehe07 /gpt-failure-cases-dataset Dataset Summary This dataset contains a curated collection of medical question–answer pairs designed to evaluate large language models (LLMs) such as GPT-4 and GPT-5 on their ability to provide factually correct responses. The dataset highlights failure cases (hallucinations) where both models struggled, making it a valuable benchmark for studying factual consistency and reliability in AI-generated medical content. Each entry consists of: question: A natural language medical query.… See the full description on the dataset page: https://huggingface.co/datasets/ehe07/gpt-failure-cases-dataset.textquestion-answering1K<n<10K0 likes42 downloads1y agoHugging Face12azizstark /synthetic-chargeback-cases Synthetic Chargeback Cases for Representment Win Prediction Dataset Description A synthetic dataset of 10,000 credit card chargeback cases designed for training and evaluating machine learning models that predict representment win probability — the likelihood a merchant will win if they fight a chargeback dispute. The dataset models realistic chargeback workflows including evidence collection, reason code categorization, customer/merchant profiling, and outcome… See the full description on the dataset page: https://huggingface.co/datasets/azizstark/synthetic-chargeback-cases.tabulartabular-classification10K<n<100K0 likes42 downloads5mo agoHugging Face13Casey27 /JailbreakPromptsgated Independent Jailbreak Datasets for LLM Guardrail Evaluation Constructed for the thesis:“Contamination Effects: How Training Data Leakage Affects Red Team Evaluation of LLM Jailbreak Detection” The effectiveness of LLM guardrails is commonly evaluated using open-source red teaming tools. However, this study reveals that significant data contamination exists between the training sets of binary jailbreak classifiers (ProtectAI, Katanemo, TestSavantAI, etc.) and the test prompts used in… See the full description on the dataset page: https://huggingface.co/datasets/Casey27/JailbreakPrompts.text1K<n<10K1 likes40 downloads6mo agoHugging Face14pbhappliedsystems /quant_eval_v7_21_per_case_results_and_run_provenance quant_eval v7.21 — Per-Case Evaluation Results and Run Provenance Supplementary evidence for the whitepaper quant_eval: A Behavioral Evaluation Harness for Full-Weight and Quantized Large Language Models. Author: Patrick Hill, PBH Applied Systems, LLC ORCID: 0009-0008-3662-1681 Licence: CC BY 4.0 Concept DOI (all versions): 10.5281/zenodo.22851375 Version DOI (this deposit): 10.5281/zenodo.22851376 What this deposit is Every quantitative result reported in the… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_v7_21_per_case_results_and_run_provenance.tabularn<1K0 likes35 downloads6d agoHugging Face15ClarusC64 /legal-case-strategy-theory-evidence-coherence-v0.1What this dataset does You receive case theory facts relied on evidence available weaknesses moves fallback You decide coherent or incoherent Daily use internal strategy checks overreach detection evidence gap flags tabulartext-classificationn<1K0 likes34 downloads7mo agoHugging Face16casecrit /2024-indonesian-electionThe dataset encompasses news articles spanning from November 29, 2023, to February 6, 2024, capturing the discourse surrounding the five presidential debates orchestrated by the General Elections Commission. Sourced from reputable platforms such as detik, kompas, and liputan6, the dataset offers a comprehensive insight into the electoral landscape and the media coverage thereof. text10K<n<100K4 likes33 downloads3y agoHugging Face17rogue-security /real-world-benign-use-cases Real-World Benign Use Cases A curated set of 178 real-world, 100%-benign examples (label == 0 for every row) pulled from production AI-coding-agent traffic — chat messages, tool output, shell commands, code snippets — built specifically to stress-test prompt-injection / jailbreak classifiers for false positives. Every row was independently judged benign with high confidence before inclusion. This is not a random sample of production traffic: rows were preferentially drawn from… See the full description on the dataset page: https://huggingface.co/datasets/rogue-security/real-world-benign-use-cases.tabulartext-classificationn<1K0 likes27 downloads2mo agoHugging Face18MichaelWittweiler /latin-intertextuality-modification-test-cases Latin Intertextuality Test Cases (Controlled Modifications) This repository provides a small qualitative dataset of Latin sentence pairs used to study how sentence-embedding similarity reacts to controlled modifications of a query sentence in intertextual links (literal quotation, paraphrase, allusion). The dataset contains 11 base cases, each paired with 12 modified variants of the query sentence (plus one random-control pairing), for a total of 143 rows. Contents… See the full description on the dataset page: https://huggingface.co/datasets/MichaelWittweiler/latin-intertextuality-modification-test-cases.textn<1K0 likes26 downloads8mo agoHugging Face19pbhappliedsystems /quant_eval_behavioral_per_case_results quant_eval — Per-case behavioral results One row per evaluation case per runner: the raw model output, every scored signal, per-case timing, expected/got pairs, the oracle trace, the fuzz audit envelope, and the decoding conditions under which the row was produced. Every aggregate statistic published in the other datasets is recomputable from this file. Part of the quant_eval public corpus: a per-case behavioral evaluation of full-weight and quantized large language models… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_behavioral_per_case_results.tabular10K<n<100K0 likes26 downloads1mo agoHugging Face20nwhite-systems /african-enterprise-ai-use-cases African Enterprise AI Use Cases Version 1.0.0 is a structured catalogue of 144 wholly synthetic use cases for responsible AI and automation planning in African operational contexts. It covers eight sectors and 24 fictional organisation archetypes. Every record defines a bounded assistance role, a named human owner, a risk category and concrete safeguards. This is a research and planning resource. The examples are not customer records, case studies or evidence that any use case… See the full description on the dataset page: https://huggingface.co/datasets/nwhite-systems/african-enterprise-ai-use-cases.textn<1K0 likes25 downloads2mo agoHugging Face21L-NLProc /Realistic_LJP_CaseSummarizertext10K<n<100K1 likes22 downloads2y agoHugging Face22ClarusC64 /legal-case-note-fact-issue-coherence-drift-v0.1What this dataset does You receive case note summary facts issues evidence references action plan risk framing You decide coherent or incoherent This mirrors daily internal file note review inside firms. tabulartext-classificationn<1K0 likes21 downloads8mo agoHugging Face23jslin09 /Fraud_Case_Verdictsgated The "Crime Facts" of "Offenses of Fraudulence" in Judicial Yuan Verdicts Dataset This data set is based on the judgments of "Offenses of Fraudulence" cases published by the Judicial Yuan. The data range of the dataset is from January 1, 2011, to December 31, 2021. 74,823 pieces of original data (judgments and rulings) were collected. We only took the contents of the "criminal facts" field of the judgment. This dataset is divided into three parts. The training dataset has 59,858… See the full description on the dataset page: https://huggingface.co/datasets/jslin09/Fraud_Case_Verdicts.texttext-generation10K<n<100K7 likes18 downloads2y agoHugging Face24CarlosKidman /test-cases Functional Test Cases This is a very small list of functional test cases that a team of software testers (QA) created for an example mobile app called Boop. Dataset Name: Boop Test Cases.csv Number of Rows: 136 Columns: 11 Test ID (int) Summary (string) Idea (string) Preconditions (string) Steps to reproduce (string) Expected Result (string) Actual Result (string) Pass/Fail (string) Bug # (string) Author (string) Area (string) 💡 There are missing values. For example… See the full description on the dataset page: https://huggingface.co/datasets/CarlosKidman/test-cases.tabularn<1K0 likes18 downloads3y agoHugging Face25vinci-grape /test_case_triggertextn<1K2 likes18 downloads3y agoHugging Face26ClarusC64 /egal-case-management-direction-order-compliance-coherence-risk-v0.1What this dataset does You receive directions order summary key deadlines diary entries task assignments compliance actions noncompliance flags You decide coherent or incoherent Daily use diary QC missed deadline risk scan ownership gaps relief planning tabulartext-classificationn<1K0 likes18 downloads7mo agoHugging Face27chriswilde006 /casetextn<1K0 likes16 downloads2y agoHugging Face28lucky-verma /aws-case-studies-and-blogs-shortThis dataset contains conversational QA pairs derived from AWS technical case studies, blogs, and documentation. Description: Content: User-assistant dialogues covering AWS services (Lambda, EC2, S3, SageMaker), architectures, and implementation scenarios from real companies like Leidos, PayEye, and Red Canary. Format: JSON messages with role (user/assistant) and content fields. Use Case: Training/fine-tuning AWS-focused chatbots or QA systems. Size: 996 entries. Key Topics: Cost… See the full description on the dataset page: https://huggingface.co/datasets/lucky-verma/aws-case-studies-and-blogs-short.textn<1K0 likes16 downloads2y agoHugging Face29HHS-Official /measles-case-and-genetic-metadata-operation-allies Measles Case and Genetic Metadata, Operation Allies Welcome Description The table contains metadata variables used to execute compartmental and genetic modeling on measles cases investigated as a component of Operation Allies Welcome. Dataset Details Publisher: Centers for Disease Control and Prevention Last Modified: 2023-07-28 Contact: Nina Masters (rhv2@cdc.gov) Source Original data can be found at: https://data.cdc.gov/d/b8tp-jsmh Usage… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/measles-case-and-genetic-metadata-operation-allies.tabularn<1K0 likes16 downloads1y agoHugging Face30HHS-Official /weekly-united-states-covid-19-cases-and-deaths-by Weekly United States COVID-19 Cases and Deaths by State - ARCHIVED Description Reporting of new Aggregate Case and Death Count data was discontinued May 11, 2023, with the expiration of the COVID-19 public health emergency declaration. This dataset will receive a final update on June 1, 2023, to reconcile historical data through May 10, 2023, and will remain publicly available. Aggregate Data Collection Process Since the start of the COVID-19 pandemic, data have been… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/weekly-united-states-covid-19-cases-and-deaths-by.tabular10K<n<100K0 likes15 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.