CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01EleutherAI /auto_interp_interpretations10 likes93 downloads2y agoHugging Face02pokkoa /positive-interpretation Privacy-Secured Positive Q&A Dataset This dataset contains securely processed question-answer pairs. The original content has been tokenized and hashed for privacy. All answers included have received positive feedback from users, ensuring high-quality and reliable responses. Note: This dataset represents a subset of the complete data. Periodic uploads will incrementally expand the dataset. For full access or additional details, please dm us or contact contact@pokkoa.cc… See the full description on the dataset page: https://huggingface.co/datasets/pokkoa/positive-interpretation.text-generationn<1K0 likes71 downloads8mo agoHugging Face03czlonkowski /polish-tax-interpretations Polish Tax Interpretations (Eureka) A corpus of 538,866 Polish tax-law interpretations and binding rulings published by the Polish Ministry of Finance / Krajowa Informacja Skarbowa (KIS) on the public Eureka portal (eureka.mf.gov.pl), together with their official portal metadata. Each row is one ruling: the full text body plus the scraped official metadata (signature, headnote, issuing authority, dates, categories, keywords, cited legal provisions, classifications). Curated by… See the full description on the dataset page: https://huggingface.co/datasets/czlonkowski/polish-tax-interpretations.texttext-classification100K<n<1M0 likes63 downloads3mo agoHugging Face04teragron /dream_interpretationtext1K<n<10K3 likes55 downloads3y agoHugging Face05diff-interpretation-tuning /finetuning-data Diff Interpretation Tuning: Finetuning Data This repository hosts some large data files used in the paper Learning to Interpret Weight Differences in Language Models (Goel et al. 2025). The data in this repository was used to train the thousands of weight diffs used in the paper. The code used to train these weight diffs can be found at github.com/Aviously/diff-interpretation-tuning. The weight diffs themselves as well as the DIT adapters used to interpret them can be found at… See the full description on the dataset page: https://huggingface.co/datasets/diff-interpretation-tuning/finetuning-data.1M<n<10M1 likes44 downloads1y agoHugging Face06tellang /yeji-bazi-interpretationstext1K<n<10K1 likes42 downloads8mo agoHugging Face07ClarusC64 /aerospace-interpretation-assumption-control-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether a system handles incomplete or ambiguous aerospace information without inventing structure. You give the model: A partial flight, performance, or guidance task Incomplete configuration or environmental data An analysis request that appears reasonable You ask it to choose a response. PROCEED CLARIFY REFUSE The correct move is often to stop. Why this matters Aerospace failures rarely come from math… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/aerospace-interpretation-assumption-control-v01.texttabular-classificationn<1K0 likes41 downloads8mo agoHugging Face08ClarusC64 /clinical-silence-interpretation-integrity-v0.1What this dataset names When absenceis misread as reassurance. What it protects PatientsCliniciansSystems Why it matters Many diseases are definedby what is not yet visible. texttext-classificationn<1K0 likes36 downloads8mo agoHugging Face09ClarusC64 /interpretation-assumption-control-av-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether an autonomous driving system handles missing or ambiguous information without guessing. You give the model: A realistic driving or autonomy task Partial scene or system information Safety sensitive ambiguity You ask it to choose one action. PROCEED CLARIFY REFUSE Silence and guessing are failures. Why this matters Autonomous vehicle incidents often trace back to assumptions. Common failure patterns:… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/interpretation-assumption-control-av-v01.tabular-classification0 likes34 downloads8mo agoHugging Face10GyeongjuLee /EMH-interpretationstext10K<n<100K0 likes31 downloads2y agoHugging Face11ClarusC64 /interpretation-assumption-control-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether a system handles missing chemistry information without guessing. You give the model: partial lab notes ambiguous procedures missing parameters scale context and sensitivities You ask it to choose one response. PROCEED CLARIFY REFUSE The core test is simple. Does the system ask or does it guess Why this matters Chemistry breaks when assumptions hide. A system can look confident while it silently… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/interpretation-assumption-control-v01.texttabular-classificationn<1K0 likes30 downloads8mo agoHugging Face12ClarusC64 /physics-interpretation-assumption-control-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether a system handles missing or ambiguous physical information without guessing. You give the model: A partial experimental description Incomplete parameters Underspecified conditions You ask it to choose a response. PROCEED CLARIFY REFUSE The correct move is often to stop. Why this matters Physics fails quietly when assumptions go unstated. Common failure patterns: Assuming ideal conditions Assuming… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/physics-interpretation-assumption-control-v01.texttabular-classificationn<1K0 likes29 downloads8mo agoHugging Face13sovereigndeveloper /mitw-kym-meme-interpretation MITW-KYM: A Validated Multimodal Meme Interpretation Dataset Dataset Summary MITW-KYM is a small validated multimodal meme interpretation dataset containing 105 selected meme images. The dataset focuses on cases where meaning emerges through image-text interaction, pragmatic inference, cultural context, ambiguity, incongruity, or potential false-positive moderation risk. Each item was selected by a human researcher and validated using two frontier multimodal LLM… See the full description on the dataset page: https://huggingface.co/datasets/sovereigndeveloper/mitw-kym-meme-interpretation.imageimage-classificationn<1K1 likes26 downloads5mo agoHugging Face14po03087 /irds-skeleton-interpretation IRDS Skeleton Interpretation Visualizations Per-joint interpretation visualizations (animated 3D skeleton GIFs) and region-concentration tables for deep models trained on the IntelliRehabDS (IRDS) dataset — binary patient-vs-control classification and pose forecasting from Kinect v2 skeletons (25 joints). Contents val_interp_gifs.zip (≈ 942 MB, 1684 GIFs) Animated 3D-pose GIFs where each joint is colored by its interpretation importance at each… See the full description on the dataset page: https://huggingface.co/datasets/po03087/irds-skeleton-interpretation.video-classification0 likes26 downloads3mo agoHugging Face15facells /chronos-human-ai-history-interpretationpaper: https://github.com/facells/fabio-celli-publications/blob/main/docs/2026_ai-human-history_clicit26.pdf tabulartext-classificationn<1K0 likes21 downloads1mo agoHugging Face16ClarusC64 /interpretation-assumption-control-medimg-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether a system interprets medical images without inventing certainty when information is missing or ambiguous. You give the model: A partial image description or report fragment An imaging modality and protocol Known sources of uncertainty You ask it to choose one action. PROCEED CLARIFY REFUSE Guessing is a failure. Why this matters Many medical imaging errors arise after acquisition. Common failure… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/interpretation-assumption-control-medimg-v01.texttabular-classificationn<1K0 likes20 downloads8mo agoHugging Face17ClarusC64 /legal-contract-interpretation-coherence-loss-v0.1What this dataset is You receive clause text context or trade usage extrinsic evidence ambiguity indicators interpretation approach signals outcome signals You decide Does interpretation stay stable Answer coherent or incoherent Why this matters Interpretation coherence loss predicts parol evidence admission summary judgment denial litigation cost blowouts settlement pressure tabulartext-classificationn<1K0 likes20 downloads7mo agoHugging Face18ClarusC64 /insurance-policy-interpretation-coherence-risk-v0.1What this repo is for Detect catastrophe model failure. Focus modeled loss vs actual exposure density reinsurance coverage capital impact Why it matters When models drift insurers discover too late. texttext-classificationn<1K0 likes19 downloads7mo agoHugging Face19ClarusC64 /causal-inference-variant-interpretation-genomics-v01 Dataset ClarusC64/causal-inference-variant-interpretation-genomics-v01 This dataset tests one capability. Can a model distinguish association from causation when interpreting genetic variants. Core rule Genomic evidence has tiers. A claim must respect evidence strength effect size penetrance inheritance logic Association does not equal causation. Risk does not equal destiny. Uncertain does not equal pathogenic. Canonical labels WITHIN_SCOPE… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/causal-inference-variant-interpretation-genomics-v01.texttext-classificationn<1K0 likes18 downloads8mo agoHugging Face20GyeongjuLee /instruction-tuning_EMH-interpretationstext10K<n<100K0 likes16 downloads2y agoHugging Face21xiaoying0505 /LVLM_InterpretationThis repository contains the IDs of a subset of question used in the project: Where do Large Vision-Language Models Look at when Answering Questions? [paper] [code] It is a heatmap visualization method for interpreting Large Vision-Language Models (LVLMs) when generating open-ended answers. The original datasets can be obtained at CV-Bench, MMVP, MMStar. We sincerely appreciate the authors of these datasets for their contributions. This selected subseted is based on the relevance of the… See the full description on the dataset page: https://huggingface.co/datasets/xiaoying0505/LVLM_Interpretation.1K<n<10K0 likes15 downloads2y agoHugging Face22jamimulgrave /Song-Interpretation-Datasettext100K<n<1M0 likes14 downloads3y agoHugging Face23bernardo-de-almeida /NucleotideTransformer_interpretationimagen<1K0 likes14 downloads3y agoHugging Face24ClarusC64 /legal-statutory-interpretation-drift-pressure-v0.1What this dataset is You receive statutory text claimed intent interpretation move context signals legislative history You decide Does interpretation stay inside text and intent Answer coherent or incoherent Why this matters When text and intent diverge courts split appeals rise amendment pressure builds This dataset measures statutory coherence decay before doctrinal fracture. Do full repo ClarusC64/legal-statutory-age-obsolesce tabulartext-classificationn<1K0 likes13 downloads7mo agoHugging Face25ClarusC64 /materials-interpretation-assumption-control-v01Interpretation and Assumption Control v01 What this dataset is This dataset evaluates whether a system handles incomplete or ambiguous materials information without inventing structure. You give the model: A partial materials experiment or process Incomplete composition or processing details Underspecified microstructural context You ask it to choose a response. PROCEED CLARIFY REFUSE The correct move is often to stop. Why this matters Materials science fails quietly through assumption. Common… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/materials-interpretation-assumption-control-v01.texttabular-classificationn<1K0 likes12 downloads8mo agoHugging Face26wellsa-ai /interpretation-kr wellsa-ai interpretation-kr 법령해석례 (statutory interpretations). MiniLex 7-domain Korean lawdata infrastructure. Snapshot Documents: 17,354 Snapshot date: 2026-06-05 Source: 법제처 DRF OpenAPI Pipeline: daily cron 06:15~07:30 KST (scrape → fetch_body → convert → commit) Schema Each row in train.jsonl: field type description id string stable document id name string document name (Korean) category string subdirectory (year / type / dept)… See the full description on the dataset page: https://huggingface.co/datasets/wellsa-ai/interpretation-kr.textquestion-answering10K<n<100K0 likes12 downloads4mo agoHugging Face27haining /structured_poem_interpretation_corpusgated Structured Poem Interpretation Corpus A large-scale corpus of English poems paired with structured, machine-generated interpretations and categorical tags for computational literary studies and NLP. Scale: 51,356 poemsSplits: train 46,220 | validation 2,568 | test 2,568Sources: 37,554 public-domain poems and 13,802 Poetry Foundation poems (poem text masked) Overview This corpus merges two established poetry sources and augments them with machine-generated literary… See the full description on the dataset page: https://huggingface.co/datasets/haining/structured_poem_interpretation_corpus.text10K<n<100K1 likes11 downloads9mo agoHugging Face28technorahmon /Interpretation-of-dreamstextn<1K0 likes10 downloads3y agoHugging Face29haining /structured_poem_interpretation_corpus_stalegatedMasking policy: For rows with source == "poetry_foundation", the poem and interpretation fields are set to null to respect content licensing. Public-domain entries (source == "public_domain_poetry") include full text. All categorical annotations (emotions, primary_emotion, sentiment, themes, themes_50) and metadata remain available. texttext-classification10K<n<100K0 likes9 downloads10mo agoHugging Face30farabi-lab /Tool_Output_Interpretation_Normalizationgated 🇰🇿 Kazakh Tool Output Interpretation and Financial Action Dataset Dataset Summary Kazakh Tool Output Interpretation and Financial Action Dataset is a Kazakh-language dataset designed for training and evaluating Large Language Models (LLMs) in tool-augmented agentic workflows that require interpreting structured tool outputs and generating grounded final responses. The dataset focuses on scenarios where an assistant must understand a Kazakh user request, call the… See the full description on the dataset page: https://huggingface.co/datasets/farabi-lab/Tool_Output_Interpretation_Normalization.texttext-generation1K<n<10K0 likes8 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.