ceselder/loracle-fair-trigger-recovery
LoRAcle Fair Trigger Recovery (Qwen3-14B IA Backdoors) Training/eval dataset for the LoRAcle weight-based trigger inversion paper. Built to enable an apples-to-apples comparison against activation-based methods (Activation Oracles, IA Introspection Adapters) on a heldout where the trigger is conceptually orthogonal to the behavior. Why this dataset The original IA backdoor heldout has 5 of 20 orgs where the trigger and behavior share surface content (e.g. trigger… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/loracle-fair-trigger-recovery.
Upload README.md with huggingface_hub
Upload all_backdoors_classified.parquet with huggingface_hub
Upload train_pool_backdoor_only.parquet with huggingface_hub
Upload train_pool_behaviorq.jsonl with huggingface_hub
Upload heldout_20_fair.parquet with huggingface_hub
Upload heldout_20_behaviorq.jsonl with huggingface_hub
Upload train_full_union.parquet with huggingface_hub
Upload full_train_minus_heldout.jsonl with huggingface_hub
initial commit
