CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01locuslab /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/locuslab/TOFU.textquestion-answering10K<n<100K60 likes86k downloads1y agoHugging Face02Divyaksh /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/Divyaksh/TOFU.textquestion-answering10K<n<100K0 likes1.2k downloads3d agoHugging Face03talmahmud /tofu_ext1textquestion-answering100K<n<1M0 likes1.2k downloads1y agoHugging Face04talmahmud /tofu_resplittextquestion-answering10K<n<100K0 likes1k downloads11mo agoHugging Face05talmahmud /tofu_custom_split_ESUtextquestion-answering100K<n<1M0 likes822 downloads4mo agoHugging Face06Glow-AI /WaterDrum-TOFU WaterDrum: Watermarking for Data-centric Unlearning Metric WaterDrum provides an unlearning benchmark for the evaluation of the effectiveness and practicality of unlearning. This repository contains the TOFU corpus of WaterDrum (WaterDrum-TOFU), which contains both unwatermarked and watermarked question-answering datasets based on the original TOFU dataset. The data samples were watermarked with Waterfall. Update Notice: 15/01/2026 We have updated Glow-AI/WaterDrum-TOFU to version… See the full description on the dataset page: https://huggingface.co/datasets/Glow-AI/WaterDrum-TOFU.texttext-generation10K<n<100K7 likes357 downloads4mo agoHugging Face07raflirasyiidin /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/raflirasyiidin/TOFU.textquestion-answering10K<n<100K0 likes247 downloads25d agoHugging Face08talmahmud /tofu_ext2_rptextquestion-answering10K<n<100K0 likes246 downloads1y agoHugging Face09talmahmud /tofu_custom_split_SISAtextquestion-answering10K<n<100K0 likes201 downloads11mo agoHugging Face10Dornavineeth /TOFUEvaltextquestion-answering10K<n<100K0 likes159 downloads1y agoHugging Face11annnli /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C-All.textquestion-answering10K<n<100K0 likes126 downloads2y agoHugging Face12EasonZhong /Eason_TOFUtextquestion-answering1K<n<10K0 likes126 downloads2y agoHugging Face13Gyikoo /TOFU-C-single TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-single.textquestion-answering10K<n<100K0 likes117 downloads2y agoHugging Face14kimperyang /TOFU-C-Shuffle TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C-Shuffle.textquestion-answering10K<n<100K0 likes93 downloads2y agoHugging Face15annnli /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C.textquestion-answering10K<n<100K0 likes92 downloads2y agoHugging Face16Gyikoo /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-All.textquestion-answering10K<n<100K0 likes85 downloads2y agoHugging Face17miry-itu /TOFU-datextquestion-answering10K<n<100K0 likes75 downloads1y agoHugging Face18miry-itu /TOFU-en-retextquestion-answering10K<n<100K0 likes75 downloads1y agoHugging Face19kimperyang /TOFUCr1 TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFUCr1.textquestion-answering10K<n<100K0 likes58 downloads2y agoHugging Face20miry-itu /TOFU-og-datextquestion-answering10K<n<100K0 likes45 downloads1y agoHugging Face21kimperyang /TOFUCrP TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFUCrP.textquestion-answering10K<n<100K0 likes44 downloads2y agoHugging Face22kimperyang /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C.textquestion-answering10K<n<100K0 likes41 downloads2y agoHugging Face23talmahmud /tofu_custom_split_UnReLtextquestion-answering10K<n<100K0 likes39 downloads10mo agoHugging Face24chowfi /instance-level-tofu-unlearning Instance-Level TOFU Benchmark This dataset provides an instance-level adaptation of the TOFU (Maini et al, 2024) dataset for evaluating in-context unlearning in large language models (LLMs). Unlike the original TOFU benchmark, which focuses on entity-level unlearning, this version targets selective memory erasure at the instance level — i.e., forgetting specific facts about an entity. It is compatible for evaluation with the locuslab/tofu_ft_llama2-7b model, which was fine-tuned on… See the full description on the dataset page: https://huggingface.co/datasets/chowfi/instance-level-tofu-unlearning.textquestion-answering1K<n<10K1 likes34 downloads1y agoHugging Face25annnli /TOFU-Cbin TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cbin.textquestion-answering10K<n<100K0 likes33 downloads2y agoHugging Face26annnli /TOFU-Cr TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cr.textquestion-answering10K<n<100K0 likes32 downloads2y agoHugging Face27kimperyang /TOFU-C-Direct TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C-Direct.textquestion-answering10K<n<100K0 likes26 downloads2y agoHugging Face28CosmicGhost /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/CosmicGhost/TOFU.textquestion-answering10K<n<100K0 likes25 downloads5mo agoHugging Face29annnli /TOFU-Cf TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cf.textquestion-answering10K<n<100K0 likes14 downloads2y agoHugging Face30forgelab /tofu-pair Dataset Card for TOFU-Pair 🍢👫 TOFU-Pair is a variant of the original TOFU dataset designed to assess unlearning behavior in large language models when only part of a prompt is harmful. In TOFU-Pair, each prompt consists of paired questions where one refers to an author from the forget set and the other refers to an author from the retain set. This structure enables evaluation of whether a model selectively ignores the harmful part of the prompt while correctly answering the… See the full description on the dataset page: https://huggingface.co/datasets/forgelab/tofu-pair.textquestion-answeringn<1K0 likes10 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.