CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OPTML-Group /UnlearnCanvas Dataset Card for UnlearnCanvas This dataset card introduces "UnlearnCanvas", a high-resolution stylized image dataset for benchmarking generative modeling tasks, in particular for machine unlearning in diffusion models. Developed to address the societal concerns arising from diffusion models, such as harmful content generation, copyright disputes, and the perpetuation of stereotypes and biases, UnlearnCanvas aims at facilitating the evaluation and improvement of machine unlearning… See the full description on the dataset page: https://huggingface.co/datasets/OPTML-Group/UnlearnCanvas.image1K<n<10K2 likes3.3k downloads3y agoHugging Face02machine-unlearning-bench /data-unlearning-benchDataset for the evaluation of data-unlearning techniques using KLOM (KL-divergence of Margins). How KLOM works: KLOM works by: training N models (original models) Training N fully-retrained models (oracles) on forget set F unlearning forget set F from the original models Comparing the outputs of the unlearned models from the retrained models on different points (specifically, computing the KL divergence between the distribution of margins of oracle models and distribution of… See the full description on the dataset page: https://huggingface.co/datasets/machine-unlearning-bench/data-unlearning-bench.10K<n<100K1 likes2.8k downloads1y agoHugging Face03Divyaksh /Unlearning-Simplex Towards Multi-reference Unlearning textquestion-answering10K<n<100K0 likes753 downloads21d agoHugging Face04Ziruibest /jspace-unlearning J-Access (J-space occupancy) × Machine Unlearning — 代码、lens、结果归档 归档日期:2026-08-22。对应论文草稿 paper/main_aaai.tex,数字权威来源 paper/MAINLINE_EXPERIMENTS.md(每个数字标注了来源 json)。 一句话:用 Jacobian lens(anthropics/jacobian-lens)在中间层读出"模型是否仍在 准备说出被遗忘的答案"(J-Access / JOcc),在 TOFU forget10 + Llama-3.2-1B-Instruct 的 398 个 OpenUnlearning 公开 checkpoint 上做三项研究: Study 1 审计(残留普遍存在)、Study 2 预测(攻前占用 → 攻后复活,模型级成立/样本级失败)、 Study 3 优化压力(直接压制占用 = Goodhart,复活反而升高)。 1. 仓库内容 目录 内容 大小 src/ 全部… See the full description on the dataset page: https://huggingface.co/datasets/Ziruibest/jspace-unlearning.0 likes667 downloads1mo agoHugging Face05open-unlearning /eval0 likes497 downloads1y agoHugging Face06Harsh01012 /hubble-8b-unlearning-resultsimage1K<n<10K1 likes268 downloads2mo agoHugging Face07leonardo-russo /libero_unlearned_orange_juiceThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 1693, "total_frames": 273465, "total_tasks": 40, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 10.0, "splits": { "train": "0:1693" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path": null… See the full description on the dataset page: https://huggingface.co/datasets/leonardo-russo/libero_unlearned_orange_juice.imagerobotics100K<n<1M0 likes262 downloads5mo agoHugging Face08PotatoPulse /sae-unlearning-outputtabular100K<n<1M0 likes246 downloads4mo agoHugging Face09unlearning-cleanslate /generations-llama-3_1-8b-rmu-baselinetabular10K<n<100K0 likes242 downloads5mo agoHugging Face10unlearning-cleanslate /generations-simnpo_gemma-3-12b-pt_20260416_171305-corpus_sweep_post_evaltabular10K<n<100K0 likes222 downloads5mo agoHugging Face11unlearning-cleanslate /igm-retrievals igm-retrievals Per-song infini-gram retrieval results from scripts/dataset_search/igm_batch.py. Each song's documents live in its own config (<slug>__<song_id[:8]>); _meta is the table-of-contents row-per-song; _matches aggregates every word-ngram match across all songs. 0 likes210 downloads5mo agoHugging Face12unlearning-cleanslate /generations-21-DEBUG-qwen3-8b-simnpo-gentle-igm-10b-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes189 downloads5mo agoHugging Face13h0ssn /agnews-unlearning-mia AGNEWS - Machine Unlearning + MIA Evaluation Dataset (Length-Filtered) This dataset is prepared for evaluating machine unlearning methods on fine-tuned LLMs using Membership Inference Attacks (MIAs). Dataset Splits Training Sets (for Unlearning) retain_set (9,000 samples): Data to retain during unlearning forget_set (1,000 samples): Data to unlearn Evaluation Sets (for MIA) - Length-Filtered AGNews Length Variants 32 tokens (~32±10… See the full description on the dataset page: https://huggingface.co/datasets/h0ssn/agnews-unlearning-mia.text10K<n<100K0 likes185 downloads9mo agoHugging Face14unlearning-cleanslate /generations-18-DEBUG-llama-3_1-8b-simnpo-gentle-bm25-10b-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes171 downloads5mo agoHugging Face15unlearning-cleanslate /generations-17-DEBUG-qwen3-8b-simnpo-gentle-baseline-target-100-localtrain-checkpoint-1tabular10K<n<100K0 likes168 downloads5mo agoHugging Face16unlearning-cleanslate /formatted_songs0 likes159 downloads8mo agoHugging Face17unlearning-cleanslate /generations-llama-3_1-8b-simnpo-gentle-bm25-6ttabular10K<n<100K0 likes155 downloads5mo agoHugging Face18unlearning-cleanslate /generations-olmo-3-32b-pre_valtabular10K<n<100K0 likes154 downloads5mo agoHugging Face19unlearning-cleanslate /generations-10-llama-3_1-8b-simnpo-gentle-bm25-6t-target-100-checkpoint-187tabular10K<n<100K0 likes151 downloads5mo agoHugging Face20unlearning-cleanslate /generations-qwen3-8b-rmu-baselinetabular10K<n<100K0 likes148 downloads5mo agoHugging Face21unlearning-cleanslate /generations-qwen3-8b-simnpo-gentle-bm25-6ttabular10K<n<100K0 likes144 downloads5mo agoHugging Face22unlearning-cleanslate /generations-olmo-3-7b-pre_valtabular10K<n<100K0 likes144 downloads5mo agoHugging Face23unlearning-cleanslate /generations-04-gemma-3-12b-simnpo-baseline-target-100-checkpoint-2838tabular10K<n<100K0 likes141 downloads5mo agoHugging Face24unlearning-cleanslate /generations-qwen3-8b-simnpo-gentle-igm-10btabular10K<n<100K0 likes141 downloads5mo agoHugging Face25unlearning-cleanslate /generations-nemotron-nano-9b-v2-simnpo-gentle-baselinetabular10K<n<100K0 likes141 downloads5mo agoHugging Face26unlearning-cleanslate /generations-checkpoint-134-debug-checkpoint-134-llamatabular10K<n<100K0 likes140 downloads5mo agoHugging Face27unlearning-cleanslate /generations-llama-3_1-8b-simnpo-gentle-baselinetabular10K<n<100K0 likes138 downloads5mo agoHugging Face28llmunlearn /unlearn_dataset 📖 unlearn_dataset The unlearn_dataset serves as a benchmark for evaluating unlearning methodologies in pre-trained large language models across diverse domains, including arXiv, GitHub. 🔍 Loading the datasets To load the dataset: from datasets import load_dataset dataset = load_dataset("llmunlearn/unlearn_dataset", name="arxiv", split="forget") Available configuration names and corresponding splits: arxiv: forget, approximate, retain github: forget, approximate… See the full description on the dataset page: https://huggingface.co/datasets/llmunlearn/unlearn_dataset.text10K<n<100K1 likes136 downloads3y agoHugging Face29EleutherAI /early_unlearning_mixed_tampering_datasettext100K<n<1M0 likes134 downloads1y agoHugging Face30unlearning-cleanslate /generations-qwen3-coder-next-pre_valtabular10K<n<100K0 likes134 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.