CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01locuslab /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/locuslab/TOFU.textquestion-answering10K<n<100K60 likes88k downloads1y agoHugging Face02talmahmud /tofu_ext1textquestion-answering100K<n<1M0 likes1.2k downloads1y agoHugging Face03talmahmud /tofu_resplittextquestion-answering10K<n<100K0 likes995 downloads11mo agoHugging Face04Divyaksh /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/Divyaksh/TOFU.textquestion-answering10K<n<100K0 likes924 downloads9h agoHugging Face05raflirasyiidin /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page… See the full description on the dataset page: https://huggingface.co/datasets/raflirasyiidin/TOFU.textquestion-answering10K<n<100K0 likes261 downloads22d agoHugging Face06talmahmud /tofu_custom_split_SISAtextquestion-answering10K<n<100K0 likes201 downloads10mo agoHugging Face07jaeunglee /uds-annotated-tofulanguage: en license: mit pretty_name: UDS-Annotated TOFU task_categories: question-answering tags: arxiv:2605.24614 unlearning llm-unlearning activation-patching tofu entity-annotation UDS-Annotated TOFU Annotated TOFU forget10 examples used in Measuring the Depth of LLM Unlearning via Activation Patching. The dataset contains factual entity and span annotations used by the Unlearning Depth Score (UDS) pipeline to evaluate whether target knowledge remains recoverable from a… See the full description on the dataset page: https://huggingface.co/datasets/jaeunglee/uds-annotated-tofu.textn<1K1 likes175 downloads2mo agoHugging Face08Dornavineeth /TOFUEvaltextquestion-answering10K<n<100K0 likes173 downloads1y agoHugging Face09annnli /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C-All.textquestion-answering10K<n<100K0 likes128 downloads2y agoHugging Face10Gyikoo /TOFU-C-single TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-single.textquestion-answering10K<n<100K0 likes125 downloads2y agoHugging Face11EasonZhong /Eason_TOFUtextquestion-answering1K<n<10K0 likes114 downloads2y agoHugging Face12Gyikoo /TOFU-C-All TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/Gyikoo/TOFU-C-All.textquestion-answering10K<n<100K0 likes102 downloads2y agoHugging Face13annnli /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-C.textquestion-answering10K<n<100K0 likes93 downloads2y agoHugging Face14kimperyang /TOFU-C-Shuffle TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C-Shuffle.textquestion-answering10K<n<100K0 likes92 downloads2y agoHugging Face15miry-itu /TOFU-datextquestion-answering10K<n<100K0 likes75 downloads1y agoHugging Face16miry-itu /TOFU-og-datextquestion-answering10K<n<100K0 likes47 downloads1y agoHugging Face17kimperyang /TOFU-C TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C.textquestion-answering10K<n<100K0 likes39 downloads2y agoHugging Face18kimperyang /TOFUCrP TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFUCrP.textquestion-answering10K<n<100K0 likes38 downloads2y agoHugging Face19annnli /TOFU-Cr TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cr.textquestion-answering10K<n<100K0 likes32 downloads2y agoHugging Face20annnli /TOFU-Cbin TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cbin.textquestion-answering10K<n<100K0 likes29 downloads2y agoHugging Face21kimperyang /TOFU-C-Direct TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/kimperyang/TOFU-C-Direct.textquestion-answering10K<n<100K0 likes27 downloads2y agoHugging Face22zekeZZ /tofu_profiletextn<1K0 likes25 downloads2y agoHugging Face23CosmicGhost /TOFU TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/CosmicGhost/TOFU.textquestion-answering10K<n<100K0 likes25 downloads5mo agoHugging Face24annnli /TOFU-Cf TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/annnli/TOFU-Cf.textquestion-answering10K<n<100K0 likes14 downloads2y agoHugging Face25Tofu0142 /camel_LongCoT Additional Information This dataset contains mathematical problem-solving traces generated using the CAMEL framework. Each entry includes: A mathematical problem statement A detailed step-by-step solution An improvement history showing how the solution was iteratively refined texttext-generationn<1K0 likes11 downloads2y agoHugging Face26eunwoneunwon /tofu-newschemetext1K<n<10K0 likes11 downloads1y agoHugging Face27forgelab /tofu-pair Dataset Card for TOFU-Pair 🍢👫 TOFU-Pair is a variant of the original TOFU dataset designed to assess unlearning behavior in large language models when only part of a prompt is harmful. In TOFU-Pair, each prompt consists of paired questions where one refers to an author from the forget set and the other refers to an author from the retain set. This structure enables evaluation of whether a model selectively ignores the harmful part of the prompt while correctly answering the… See the full description on the dataset page: https://huggingface.co/datasets/forgelab/tofu-pair.textquestion-answeringn<1K0 likes10 downloads2y agoHugging Face28KaivuH /tofu_original_copy_finaltext1K<n<10K0 likes3 downloads2y agoHugging Face29Novaspree /tofu-gemma3-adapter-negation-predictionstext1K<n<10K0 likes2 downloads4mo agoHugging Face30KaivuH /tofu_wikipedia_finaltest textn<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.