CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ssadqdsacf /cross-unlearning-case-400imagen<1K0 likes132 downloads3d agoHugging Face02boyiwei /copyright_unlearningtextquestion-answering1K<n<10K0 likes119 downloads2y agoHugging Face03shichenghu /personal-info-unlearning Synthetic Personal Information Unlearning Dataset Dataset Description This dataset is designed for research on large language model (LLM) unlearning in controlled synthetic personal-information settings. It contains synthetic profiles and question-answer data for four personal attributes: Year of birth Blood type Postcode Social insurance number The benchmark provides three forget-set sizes: N = 5, 20, 40. All personal-profile data are synthetically generated… See the full description on the dataset page: https://huggingface.co/datasets/shichenghu/personal-info-unlearning.textquestion-answering100K<n<1M0 likes94 downloads25d agoHugging Face04enver /classical-arabic-logic-slop-unlearning 📜 Classical Arabic Logic & Code Slop Unlearning Dataset Epistemic Alignment & Anti-Pattern Elimination for Sovereign Code Synthesis Author: Enver AynEngineAffiliation: Sovereign Epistemic AI Research / AynEngine Project (Switzerland)Associated Paper: Hierarchical Symbolic-Neural Mixture of Experts (H-MoE) 🏛️ Dataset Overview Modern large language models trained on massive internet web crawls frequently hallucinate architectural anti-patterns… See the full description on the dataset page: https://huggingface.co/datasets/enver/classical-arabic-logic-slop-unlearning.textn<1K0 likes49 downloads13d agoHugging Face05LZ12DH /unlearning TOFU: Task of Fictitious Unlearning 🍢 The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks. The dataset comprises question-answer pairs based on autobiographies of 200 different authors that do not exist and are completely fictitiously generated by the GPT-4 model. The goal of the task is to unlearn a fine-tuned model on various fractions of the forget set. Quick Links Website: The landing page for TOFU… See the full description on the dataset page: https://huggingface.co/datasets/LZ12DH/unlearning.textquestion-answering1K<n<10K0 likes44 downloads2y agoHugging Face06EmanuelBP /syntetic-adversarial-unlearning-V1text100K<n<1M0 likes4 downloads3mo agoHugging Face07YipingZhang /unlearning_negative_jsonltextn<1K0 likes3 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.