datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
2025-challenge-task-instances2025-challenge-task-instancesvn-provinces-criminal-cases-first-instance
Vietnam criminal cases first-instance trial
Vietnam criminal cases first-instance trial. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (189 rows)
data/provinces.csv
data/provinces.dta
data/provinces.xlsx
regions (18 rows)
data/regions.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-criminal-cases-first-instance.instance-level-tofu-unlearning
Instance-Level TOFU Benchmark
This dataset provides an instance-level adaptation of the TOFU (Maini et al, 2024) dataset for evaluating in-context unlearning in large language models (LLMs). Unlike the original TOFU benchmark, which focuses on entity-level unlearning, this version targets selective memory erasure at the instance level — i.e., forgetting specific facts about an entity.
It is compatible for evaluation with the locuslab/tofu_ft_llama2-7b model, which was fine-tuned on… See the full description on the dataset page: https://huggingface.co/datasets/chowfi/instance-level-tofu-unlearning.magic_gamma_1500_instancesadult_1500_instancesLaMini-instruction-5k_Instances
