datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
data-unlearning-benchDataset for the evaluation of data-unlearning techniques using KLOM (KL-divergence of Margins).
How KLOM works:
KLOM works by:
training N models (original models)
Training N fully-retrained models (oracles) on forget set F
unlearning forget set F from the original models
Comparing the outputs of the unlearned models from the retrained models on different points
(specifically, computing the KL divergence between the distribution of margins of oracle models and distribution of… See the full description on the dataset page: https://huggingface.co/datasets/machine-unlearning-bench/data-unlearning-bench.hubble-8b-unlearning-resultsjspace-unlearning
J-Access (J-space occupancy) × Machine Unlearning — 代码、lens、结果归档
归档日期:2026-08-22。对应论文草稿 paper/main_aaai.tex,数字权威来源
paper/MAINLINE_EXPERIMENTS.md(每个数字标注了来源 json)。
一句话:用 Jacobian lens(anthropics/jacobian-lens)在中间层读出"模型是否仍在
准备说出被遗忘的答案"(J-Access / JOcc),在 TOFU forget10 + Llama-3.2-1B-Instruct
的 398 个 OpenUnlearning 公开 checkpoint 上做三项研究:
Study 1 审计(残留普遍存在)、Study 2 预测(攻前占用 → 攻后复活,模型级成立/样本级失败)、
Study 3 优化压力(直接压制占用 = Goodhart,复活反而升高)。
1. 仓库内容
目录
内容
大小
src/
全部… See the full description on the dataset page: https://huggingface.co/datasets/Ziruibest/jspace-unlearning.Unlearning-Simplex
Towards Multi-reference Unlearning
evalsae-unlearning-outputLKF-unlearning_Jewell_new_config_no_filter_rephrasings_finalLKF-unlearning_Jewell_new_config_no_filter_finalgenerations-simnpo_gemma-3-12b-pt_20260416_171305-corpus_sweep_post_evalgenerations-llama-3_1-8b-rmu-baselineigm-retrievals
igm-retrievals
Per-song infini-gram retrieval results from scripts/dataset_search/igm_batch.py. Each song's documents live in its own config (<slug>__<song_id[:8]>); _meta is the table-of-contents row-per-song; _matches aggregates every word-ngram match across all songs.
generations-21-DEBUG-qwen3-8b-simnpo-gentle-igm-10b-target-100-localtrain-checkpoint-1generations-17-DEBUG-qwen3-8b-simnpo-gentle-baseline-target-100-localtrain-checkpoint-1generations-18-DEBUG-llama-3_1-8b-simnpo-gentle-bm25-10b-target-100-localtrain-checkpoint-1generations-qwen3-8b-rmu-baselinegenerations-qwen3-8b-simnpo-gentle-bm25-6tgenerations-14-llama-3_1-8b-rmu-baseline-target-100-checkpoint-1722formatted_songsgenerations-olmo-3-32b-pre_valgenerations-10-llama-3_1-8b-simnpo-gentle-bm25-6t-target-100-checkpoint-187generations-llama-3_1-8b-simnpo-gentle-baselinegenerations-llama-3_1-8b-simnpo-gentle-bm25-6tagnews-unlearning-mia
AGNEWS - Machine Unlearning + MIA Evaluation Dataset (Length-Filtered)
This dataset is prepared for evaluating machine unlearning methods on fine-tuned LLMs using Membership Inference Attacks (MIAs).
Dataset Splits
Training Sets (for Unlearning)
retain_set (9,000 samples): Data to retain during unlearning
forget_set (1,000 samples): Data to unlearn
Evaluation Sets (for MIA) - Length-Filtered
AGNews Length Variants
32 tokens (~32±10… See the full description on the dataset page: https://huggingface.co/datasets/h0ssn/agnews-unlearning-mia.generations-checkpoint-134-debug-checkpoint-134-llamagenerations-04-gemma-3-12b-simnpo-baseline-target-100-checkpoint-2838generations-nemotron-nano-9b-v2-simnpo-gentle-baselinegenerations-olmo-3-7b-pre_valgenerations-qwen3-8b-simnpo-gentle-igm-10bgenerations-qwen3-coder-next-pre_valgenerations-qwen3-8b-undial-baseline
