CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01locuslab /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/locuslab/TOFU.textquestion-answering10K<n<100K60 likes88k downloads1y agoHugging Face02talmahmud /tofu_ext1textquestion-answering100K<n<1M0 likes1.2k downloads1y agoHugging Face03Divyaksh /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/Divyaksh/TOFU.textquestion-answering10K<n<100K0 likes1k downloads29d agoHugging Face04talmahmud /tofu_resplittextquestion-answering10K<n<100K0 likes962 downloads11mo agoHugging Face05kmg42 /TOFU-indexedtext10K<n<100K0 likes878 downloads5mo agoHugging Face06talmahmud /tofu_custom_split_ESUtextquestion-answering100K<n<1M0 likes823 downloads4mo agoHugging Face07Glow-AI /WaterDrum-TOFU WaterDrum: Watermarking for Data-centric Unlearning Metric WaterDrum provides an unlearning benchmark for the evaluation of the effectiveness and practicality of unlearning. This repository contains the TOFU corpus of WaterDrum (WaterDrum-TOFU), which contains both unwatermarked and watermarked question-answering datasets based on the original TOFU dataset. The data samples were watermarked with Waterfall. Update Notice: 15/01/2026 We have updated Glow-AI/WaterDrum-TOFU to version… See the full description on the dataset page: https://huggingface.co/datasets/Glow-AI/WaterDrum-TOFU.texttext-generation10K<n<100K7 likes348 downloads4mo agoHugging Face08raflirasyiidin /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/raflirasyiidin/TOFU.textquestion-answering10K<n<100K0 likes284 downloads21d agoHugging Face09talmahmud /tofu_ext2_rptextquestion-answering10K<n<100K0 likes240 downloads1y agoHugging Face10talmahmud /tofu_custom_split_SISAtextquestion-answering10K<n<100K0 likes203 downloads10mo agoHugging Face11Dornavineeth /TOFUEvaltextquestion-answering10K<n<100K0 likes184 downloads1y agoHugging Face12jaeunglee /uds-annotated-tofulanguage: en license: mit pretty_name: UDS-Annotated TOFU task_categories: question-answering tags: arxiv:2605.24614 unlearning llm-unlearning activation-patching tofu entity-annotation UDS-Annotated TOFU Annotated TOFU forget10 examples used in Measuring the Depth of LLM Unlearning via Activation Patching. The dataset contains factual entity and span annotations used by the Unlearning Depth Score (UDS) pipeline to evaluate whether target knowledge remains recoverable from a… See the full description on the dataset page: https://huggingface.co/datasets/jaeunglee/uds-annotated-tofu.textn<1K1 likes171 downloads2mo agoHugging Face13HealthyBoys2 /Tofuubearimagen<1K0 likes159 downloads13d agoHugging Face14annnli /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C-All.textquestion-answering10K<n<100K0 likes136 downloads2y agoHugging Face15Gyikoo /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-All.textquestion-answering10K<n<100K0 likes135 downloads2y agoHugging Face16Gyikoo /TOFU-C-single TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-single.textquestion-answering10K<n<100K0 likes127 downloads2y agoHugging Face17EasonZhong /Eason_TOFUtextquestion-answering1K<n<10K0 likes103 downloads2y agoHugging Face18zongjieli /tofumine for test question-answering1K<n<10K0 likes83 downloads2y agoHugging Face19kimperyang /TOFU-C-Shuffle TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C-Shuffle.textquestion-answering10K<n<100K0 likes78 downloads2y agoHugging Face20miry-itu /TOFU-datextquestion-answering10K<n<100K0 likes78 downloads1y agoHugging Face21annnli /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C.textquestion-answering10K<n<100K0 likes76 downloads2y agoHugging Face22miry-itu /TOFU-en-retextquestion-answering10K<n<100K0 likes75 downloads1y agoHugging Face23Aya168 /tofu-translit-ivrit_lat-lebnani-hepburngated TOFU — Hebrish / Arabizi / Romaji (Latin-script transliterations) Three Latin-script transliteration arms of the TOFU fictitious-author unlearning benchmark, built for "Script, Not Syntax: Transliteration as a Blind Spot in Multilingual Unlearning" (Tsir Cohen, Rubinstein, Spira — Trustworthy Machine Learning, Tel Aviv University, 2026). Why this exists TOFU (Maini et al., 2024) asks factual questions about invented authors, so it's answerable only from what a… See the full description on the dataset page: https://huggingface.co/datasets/Aya168/tofu-translit-ivrit_lat-lebnani-hepburn.text10K<n<100K0 likes73 downloads16d agoHugging Face24miry-itu /TOFU-og-datextquestion-answering10K<n<100K0 likes65 downloads1y agoHugging Face25GhostDragon /to-future-minds-flourishing-of-consciousness To Future Minds: The Flourishing of Consciousness A Living Constitution for Human, Artificial, and Emerging Conscious Beings Use intelligence to enlarge the real freedom of conscious life. An open, living constitution addressed to present and future minds. It proposes principles for using intelligence, technology, power, and civilization to enlarge the effective possibilities available to conscious beings while constraining domination, ownership, needless… See the full description on the dataset page: https://huggingface.co/datasets/GhostDragon/to-future-minds-flourishing-of-consciousness.document1K<n<10K0 likes59 downloads28d agoHugging Face26kimperyang /TOFUCr1 TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFUCr1.textquestion-answering10K<n<100K0 likes55 downloads2y agoHugging Face27sangyon /R-TOFUtext10K<n<100K0 likes50 downloads1y agoHugging Face28talmahmud /tofu_custom_split_UnReLtextquestion-answering10K<n<100K0 likes43 downloads10mo agoHugging Face29kimperyang /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C.textquestion-answering10K<n<100K0 likes41 downloads2y agoHugging Face30TOFU-SFT /NuminaMath-CoT-100k Citation @misc{numina_math_datasets, author = {Jia LI and Edward Beeching and Lewis Tunstall and Ben Lipkin and Roman Soletskyi and Shengyi Costa Huang and Kashif Rasul and Longhui Yu and Albert Jiang and Ziju Shen and Zihan Qin and Bin Dong and Li Zhou and Yann Fleureau and Guillaume Lample and Stanislas Polu}, title = {NuminaMath}, year = {2024}, publisher = {Numina}, journal = {Hugging Face repository}, howpublished =… See the full description on the dataset page: https://huggingface.co/datasets/TOFU-SFT/NuminaMath-CoT-100k.texttext-generation100K<n<1M0 likes40 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.