CoolFace
Datasetpublic

spadeMIA/GoodWiki_Corpus_1024_2040

GoodWiki 1024–2040: paragraph-truncated MIA fine-tuning corpus A deterministic, paragraph-truncated corpus of English Wikipedia Good/Featured articles, built from euirim/goodwiki for membership inference attack (MIA) experiments on fine-tuned language models. Membership labels are defined relative to the fine-tuning population. train (10,000 rows) is the only split used for fine-tuning, and every row has label = 1. test (1,000 rows) remains held out, and every row has label =… See the full description on the dataset page: https://huggingface.co/datasets/spadeMIA/GoodWiki_Corpus_1024_2040.

sourceHugging Facecc-by-sa-4.0updated 2mo agoView on Hugging Face
0likes273downloads
settings

This repository belongs to spadeMIA on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameGoodWiki_Corpus_1024_2040
visibilitypublic
licencecc-by-sa-4.0
gatedno
ownerspadeMIA
Account settings
spadeMIA/GoodWiki_Corpus_1024_2040 · CoolFace