CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OPTML-Group /UnlearnCanvas Dataset Card for UnlearnCanvas This dataset card introduces "UnlearnCanvas", a high-resolution stylized image dataset for benchmarking generative modeling tasks, in particular for machine unlearning in diffusion models. Developed to address the societal concerns arising from diffusion models, such as harmful content generation, copyright disputes, and the perpetuation of stereotypes and biases, UnlearnCanvas aims at facilitating the evaluation and improvement of machine unlearning… See the full description on the dataset page: https://huggingface.co/datasets/OPTML-Group/UnlearnCanvas.image1K<n<10K2 likes3.3k downloads3y agoHugging Face02Divyaksh /Unlearning-Simplex Towards Multi-reference Unlearning textquestion-answering10K<n<100K0 likes753 downloads21d agoHugging Face03unlearning-cleanslate /generations-llama-3_1-8b-rmu-baselinetabular10K<n<100K0 likes242 downloads5mo agoHugging Face04unlearning-cleanslate /generations-simnpo_gemma-3-12b-pt_20260416_171305-corpus_sweep_post_evaltabular10K<n<100K0 likes222 downloads5mo agoHugging Face05unlearning-cleanslate /generations-21-DEBUG-qwen3-8b-simnpo-gentle-igm-10b-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes189 downloads5mo agoHugging Face06h0ssn /agnews-unlearning-mia AGNEWS - Machine Unlearning + MIA Evaluation Dataset (Length-Filtered) This dataset is prepared for evaluating machine unlearning methods on fine-tuned LLMs using Membership Inference Attacks (MIAs). Dataset Splits Training Sets (for Unlearning) retain_set (9,000 samples): Data to retain during unlearning forget_set (1,000 samples): Data to unlearn Evaluation Sets (for MIA) - Length-Filtered AGNews Length Variants 32 tokens (~32±10… See the full description on the dataset page: https://huggingface.co/datasets/h0ssn/agnews-unlearning-mia.text10K<n<100K0 likes185 downloads9mo agoHugging Face07unlearning-cleanslate /generations-18-DEBUG-llama-3_1-8b-simnpo-gentle-bm25-10b-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes171 downloads5mo agoHugging Face08unlearning-cleanslate /generations-17-DEBUG-qwen3-8b-simnpo-gentle-baseline-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes168 downloads5mo agoHugging Face09unlearning-cleanslate /generations-llama-3_1-8b-simnpo-gentle-bm25-6ttabular10K<n<100K0 likes155 downloads5mo agoHugging Face10unlearning-cleanslate /generations-olmo-3-32b-pre_valtabular10K<n<100K0 likes154 downloads5mo agoHugging Face11unlearning-cleanslate /generations-10-llama-3_1-8b-simnpo-gentle-bm25-6t-target-100-checkpoint-187tabular10K<n<100K0 likes151 downloads5mo agoHugging Face12unlearning-cleanslate /generations-qwen3-8b-rmu-baselinetabular10K<n<100K0 likes148 downloads5mo agoHugging Face13unlearning-cleanslate /generations-qwen3-8b-simnpo-gentle-bm25-6ttabular10K<n<100K0 likes144 downloads5mo agoHugging Face14unlearning-cleanslate /generations-olmo-3-7b-pre_valtabular10K<n<100K0 likes144 downloads5mo agoHugging Face15unlearning-cleanslate /generations-04-gemma-3-12b-simnpo-baseline-target-100-checkpoint-2838tabular10K<n<100K0 likes141 downloads5mo agoHugging Face16unlearning-cleanslate /generations-qwen3-8b-simnpo-gentle-igm-10btabular10K<n<100K0 likes141 downloads5mo agoHugging Face17unlearning-cleanslate /generations-nemotron-nano-9b-v2-simnpo-gentle-baselinetabular10K<n<100K0 likes141 downloads5mo agoHugging Face18unlearning-cleanslate /generations-checkpoint-134-debug-checkpoint-134-llamatabular10K<n<100K0 likes140 downloads5mo agoHugging Face19unlearning-cleanslate /generations-llama-3_1-8b-simnpo-gentle-baselinetabular10K<n<100K0 likes138 downloads5mo agoHugging Face20llmunlearn /unlearn_dataset 📖 unlearn_dataset The unlearn_dataset serves as a benchmark for evaluating unlearning methodologies in pre-trained large language models across diverse domains, including arXiv, GitHub. 🔍 Loading the datasets To load the dataset: from datasets import load_dataset dataset = load_dataset("llmunlearn/unlearn_dataset", name="arxiv", split="forget") Available configuration names and corresponding splits: arxiv: forget, approximate, retain github: forget, approximate… See the full description on the dataset page: https://huggingface.co/datasets/llmunlearn/unlearn_dataset.text10K<n<100K1 likes136 downloads3y agoHugging Face21EleutherAI /early_unlearning_mixed_tampering_datasettext100K<n<1M0 likes134 downloads1y agoHugging Face22unlearning-cleanslate /generations-qwen3-coder-next-pre_valtabular10K<n<100K0 likes134 downloads5mo agoHugging Face23unlearning-cleanslate /generations-14-llama-3_1-8b-rmu-baseline-target-100-checkpoint-1722tabular10K<n<100K0 likes134 downloads5mo agoHugging Face24ssadqdsacf /cross-unlearning-case-400imagen<1K0 likes133 downloads6d agoHugging Face25unlearning-cleanslate /generations-qwen3-8b-undial-baselinetabular10K<n<100K0 likes132 downloads5mo agoHugging Face26unlearning-cleanslate /generations-03-gemma-3-12b-simnpo-gentle-baseline-target-100-checkpoint-1419tabular10K<n<100K0 likes130 downloads5mo agoHugging Face27unlearning-cleanslate /generations-16-DEBUG-llama-3_1-8b-simnpo-gentle-baseline-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes129 downloads5mo agoHugging Face28unlearning-cleanslate /generations-nemotron-nano-9b-v2-simnpo-baselinetabular10K<n<100K0 likes127 downloads5mo agoHugging Face29unlearning-cleanslate /generations-15-qwen3-8b-rmu-baseline-target-100-checkpoint-1078tabular10K<n<100K0 likes126 downloads5mo agoHugging Face30boyiwei /copyright_unlearningtextquestion-answering1K<n<10K0 likes124 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.