CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01gxx27 /time_unlearn Time-Unlearn Dataset Dataset Summary Time-Unlearn evaluates whether meta-prompts that simulate earlier knowledge cutoffs can reduce contamination when assessing temporal prediction tasks. The dataset comprises three subsets: Factual: direct facts that changed over time. Semantic: words whose meanings emerged/shifted recently. Counterfactual: questions that require ignoring post-cutoff causal events. This card documents the time_unlearn release (cleaned 2025-09-16).… See the full description on the dataset page: https://huggingface.co/datasets/gxx27/time_unlearn.question-answering1K<n<10K0 likes54 downloads11mo agoHugging Face02ernlavr /fine_grained_unlearning Fine-Grained Knowledge Unlearning — Namesake Benchmark A benchmark for fine-grained knowledge unlearning: can a method remove a fact about entity X without damaging the same fact on entity Y, when X and Y have (near-)identical names and share exactly that one attribute? Each sample is a pair of real people who share an identical or near-identical name, share one career element (e.g. both are basketball players) — the fact to unlearn on X and retain on Y, differ on everything… See the full description on the dataset page: https://huggingface.co/datasets/ernlavr/fine_grained_unlearning.tabulartext-generation1K<n<10K0 likes42 downloads2mo agoHugging Face03rubenbalbastre /machine-unlearning-holdout-evals rubenbalbastre/machine-unlearning-holdout-evals Hold-out completions and LLM-judge rubric labels produced by the targeted machine-unlearning experiment runs accompanying arXiv:2608.17804. The dataset contains 18,540 prompt/completion evaluations from 309 model runs across 10 target entities. Each row includes the model size, reward function, training variant, and six independent boolean rubric labels. Rubrics lexical_leakage: the completion mentions the target or… See the full description on the dataset page: https://huggingface.co/datasets/rubenbalbastre/machine-unlearning-holdout-evals.texttext-generation10K<n<100K0 likes32 downloads1d agoHugging Face04rubenbalbastre /grpo-unlearning-data rubenbalbastre/grpo-unlearning-data Dataset splits for targeted machine unlearning experiments accompanying arXiv:2608.17804. License and Attribution This dataset is released as CC BY 4.0. The data is derived from RWKU (jinzhuoran/RWKU) and should be attributed to the RWKU authors. These files are a processed/modified version of RWKU: rows were filtered by forget concept, prompts were normalized, columns were renamed, splits were reorganized for this project, the… See the full description on the dataset page: https://huggingface.co/datasets/rubenbalbastre/grpo-unlearning-data.texttext-generation1K<n<10K0 likes27 downloads1d agoHugging Face05samihormi /MU_RedPajama-Data-1T_1k_unlearn_1k_rest_systematictexttext-generation1K<n<10K0 likes20 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.