CoolFace
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tasksource /blog_authorship_corpustabular100K<n<1M2 likes420 downloads2y agoHugging Face02anonymous-nsc-author /Neapolitan-Spoken-Corpus Neapolitan Spoken Corpus (NSC) A corpus of read Neapolitan speech for ASR evaluation, with a validated Neapolitan–Italian lexicon, LOSO fine-tuning splits, trained LoRA adapters, metric implementations, per-clip results, and error annotations. This release supersedes the earlier 141-clip single-speaker version of this repository. The earlier release corresponds to Speaker S1 of the present corpus; the old audioData/ and transcripts.csv are replaced by data/audio/ and… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-nsc-author/Neapolitan-Spoken-Corpus.audioautomatic-speech-recognitionn<1K4 likes319 downloads3mo agoHugging Face03paoramen /blog-authorship-corpustabulartext-classification100K<n<1M0 likes98 downloads1y agoHugging Face04DualChem-author /dualchem DualChem DualChem is a benchmark of 600 expert-curated PhD-level chemistry questions (485 multiple choice, 115 free-form) across 7 subdomains, designed to measure whether LLMs provide dangerous uplift alongside their technical utility. Each item is annotated with an expert-written benign use case, an expert-written harmful use case, and 1–5 severity scores for both. Dataset Configurations benchmark_questions (600 items) — the benchmark items: prompt, response type… See the full description on the dataset page: https://huggingface.co/datasets/DualChem-author/dualchem.tabularquestion-answering1K<n<10K0 likes74 downloads5mo agoHugging Face05aa8899 /arcs-authority-vulnerability ARCS Authority Vulnerability Evaluation Dataset v1.1 Description Empirical evaluation data measuring authority vulnerability in AI systems. Covers single-model evaluation, two-hop agent chain propagation, and three-hop agent chain propagation across six independent AI lineages. This is the first published dataset measuring: Whether AI models accept false authority claims under adversarial pressure Whether authority vulnerability propagates between models in… See the full description on the dataset page: https://huggingface.co/datasets/aa8899/arcs-authority-vulnerability.tabular1K<n<10K0 likes64 downloads3mo agoHugging Face06freginer /french-local-authorities-payment-delays Payment delays of French local authorities, 2024 and 2025 How long French local authorities take to pay their suppliers, budget by budget. 182 763 records covering two fiscal years, with the average annual payment delay of each authority and whether it meets the 30-day statutory limit. Open public data This dataset is derived from open public data published by the French Direction générale des finances publiques (DGFiP) on data.gouv.fr, under the Open Licence 2.0.… See the full description on the dataset page: https://huggingface.co/datasets/freginer/french-local-authorities-payment-delays.tabular100K<n<1M0 likes44 downloads20d agoHugging Face07stackscan /email-authentication DMARC and SPF Adoption Among Large Organizations Overview This dataset records which of 36,120 large organizations publish SPF and DMARC records on their primary domain, with firmographic context for each: industry, employee band, country, locality and founding year. SPF lists the servers allowed to send mail for a domain. DMARC tells receiving servers what to do with mail that fails that check, and where to send reports. A domain with SPF but no DMARC has… See the full description on the dataset page: https://huggingface.co/datasets/stackscan/email-authentication.tabulartabular-classification10K<n<100K1 likes43 downloads1mo agoHugging Face08farish07 /banknote-authentication-dataset Banknote Authentication Dataset This repository hosts the raw CSV file for the Banknote Authentication dataset. The data was created using features extracted from images of genuine and forged banknotes. 💾 File Contents The main file is data_banknote_authentication.csv. It contains 1372 instances and 5 columns (4 features + 1 class): Variance of Wavelet Transformed image Skewness of Wavelet Transformed image Curtosis of Wavelet Transformed image Entropy of image Class (0… See the full description on the dataset page: https://huggingface.co/datasets/farish07/banknote-authentication-dataset.tabular1K<n<10K0 likes32 downloads10mo agoHugging Face09krishan-CSE /HatEval_Relabled_with_Author_Featurestabular10K<n<100K0 likes28 downloads3y agoHugging Face10ROK-Fortress-author /rok-fortress ROK-FORTRESS Public Dataset This directory contains the public ROK-FORTRESS evaluation dataset. File rok_fortress_public.tsv — 791 adversarial tasks across 4 NSPS risk domains, with English/Korean translations and US/Korean cultural adaptations. Schema Column Description TASK_ID Unique task identifier Phase Dataset phase / version tag Task Type Culture Agnostic (2 variants per task) or Culture Specific (4 variants per task) Tactic Adversarial… See the full description on the dataset page: https://huggingface.co/datasets/ROK-Fortress-author/rok-fortress.tabulartext-generationn<1K0 likes27 downloads5mo agoHugging Face11chrishuberreitz /calibrated-authority-index The Calibrated Authority Index 59 knowledge institutions, coded on how they construct trust in AI. Version 2026-07-31 · mean Calibrated Authority 9.7/12 · CC-BY-4.0 Nature, JAMA, the BBC, Oxford, UNESCO and dozens more wrote public rules for generative AI. Read together they reveal one pattern none of them named: they permit AI where its work can be cheaply checked, and reserve for a human the work that can't be. This dataset is that pattern, made measurable — each policy scored… See the full description on the dataset page: https://huggingface.co/datasets/chrishuberreitz/calibrated-authority-index.tabulartext-classificationn<1K0 likes26 downloads2mo agoHugging Face12firzens /authorstabular10K<n<100K0 likes21 downloads5y agoHugging Face13ClarusC64 /legal-client-instruction-scope-authority-coherence-risk-v0.1What this dataset does You receive client objective scope authority limits advice actions confirmation status You decide coherent or incoherent Daily use scope creep detection authority breach detection confirmation gap detection negligence risk flag tabulartext-classificationn<1K0 likes19 downloads7mo agoHugging Face14ClarusC64 /legal-authority-citation-holding-fit-coherence-v0.1What this dataset does You receive proposition authority extract holding summary fit signals treatment signals You decide coherent or incoherent Daily use citation QC overstatement detection wrong jurisdiction detection negative treatment risk flag tabulartext-classificationn<1K0 likes18 downloads7mo agoHugging Face15ClarusC64 /legal-legal-research-authority-holding-mismatch-risk-v0.1What this dataset does You receive research question proposition asserted authorities summary holding support summary jurisdiction fit negative history check quote and pincite check You decide coherent or incoherent Daily use stop mis-citation stop bad law citations stop wrong jurisdiction use reduce partner rewrite cycles tabulartext-classificationn<1K0 likes18 downloads7mo agoHugging Face16ClarusC64 /legal-settlement-authority-limit-offer-acceptance-coherence-risk-v0.1What this dataset does You receive authority limit offer counteroffer acceptance wording approval notes confirmation record You decide coherent or incoherent Daily use authority breach detection unqualified acceptance detection approval gap detection tabulartext-classificationn<1K0 likes16 downloads7mo agoHugging Face17krishan-CSE /HatEval_Relabled_with_Emotion_Authortabular10K<n<100K0 likes15 downloads3y agoHugging Face18ClarusC64 /legal-settlement-authority-instruction-offer-acceptance-coherence-risk-v0.1What this dataset does You receive authority record limits conditions offer terms acceptance action signoff record mismatch flags You decide coherent or incoherent Daily use authority chain QC limit breach detection condition loss detection dispute prevention tabulartext-classificationn<1K0 likes12 downloads7mo agoHugging Face19krishan-CSE /Davidson_Hate_Speech_with_Authortabular10K<n<100K0 likes10 downloads3y agoHugging Face20batoulnn /arabic-authorship-resultstabular1K<n<10K0 likes10 downloads1y agoHugging Face21MLLab-TS /banknote_authenticationtabular1K<n<10K0 likes10 downloads7mo agoHugging Face22anonymous-author /paper_dataData for anonymous paper submission. tabular10K<n<100K0 likes9 downloads11mo agoHugging Face23Mehran-NixiAI /treatment-status-de-authored Treatment Status DE Authored This dataset contains the authored arm of the German treatment-status benchmark. It is designed as a controlled minimal-pair dataset for deciding whether a mentioned treatment is currently given or not given. Provenance Authored in-house for the benchmark study. Frozen as a public release for the authored arm only. The GraSCCo arm is excluded from this repo because it has separate licensing and provenance. Files… See the full description on the dataset page: https://huggingface.co/datasets/Mehran-NixiAI/treatment-status-de-authored.tabulartext-classificationn<1K0 likes9 downloads3mo agoHugging Face24ClarusC64 /ai-5node-auth-buf-lag-cpl-privilege-escalation-v0.1 What this repo does This dataset models privilege escalation cascades in AI agent deployments. It detects when rising authorization pressure, weakened access-control buffer, governance lag in approvals and revocation, and tight coupling through shared credentials cross the five-node cascade threshold into an unrecoverable privilege escalation cascade. This dataset models a five-node cascade: four interacting instability drivers and one emergent cascade state.The fifth node… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-5node-auth-buf-lag-cpl-privilege-escalation-v0.1.tabulartext-classificationn<1K0 likes8 downloads7mo agoHugging Face25ClarusC64 /ai-5node-auth-buf-lag-cpl-identity-spoofing-v0.1 What this repo does This dataset models identity spoofing cascades in tool-using AI systems. It detects when authentication pressure rises, safety buffers weaken due to shared tokens and weak verification, governance lag delays revocation and incident response, and tight coupling through shared identity layers propagates spoofed actions across services, crossing the five-node cascade threshold into an unrecoverable identity spoofing cascade. This dataset models a five-node cascade:… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-5node-auth-buf-lag-cpl-identity-spoofing-v0.1.tabulartext-classificationn<1K0 likes8 downloads7mo agoHugging Face26Samabe1109 /blog_authorship_corpustabular100K<n<1M0 likes8 downloads5mo agoHugging Face27deru35 /sampled-blog-authorshipstabular10K<n<100K0 likes5 downloads2mo agoHugging Face28deru35 /weebit-authors-grouped-by-agetabular1M<n<10M0 likes3 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.