CoolFace
Datasetpublic

disentangled-vla/libero-island-ablation-v12-sigreg-medium

V12 — Single head + SIGReg medium weight (λ=0.01) Bridges V3 (λ_sig=0.1, sig-dominated) and V8 (λ_sig=0.00025, action-dominated). Architecture / Training Backbone: Qwen2.5-VL-3B-Instruct + LoRA (r=32, q/v_proj) Single AttentiveLatentHead A + ResNet action head Loss: 1.0·L1_action + 0.01·SIGReg(z_a, z_domain_current) 100K steps, bs=64, lr=1e-4 cosine, 10 epochs Results — epoch_10 Setting SR AA 0.72 BB 0.72 C_A 0.12 C_B 0.24 AL_B… See the full description on the dataset page: https://huggingface.co/datasets/disentangled-vla/libero-island-ablation-v12-sigreg-medium.

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
0likes315downloads
Dataset Card

V12 — Single head + SIGReg medium weight (λ=0.01)

Bridges V3 (λsig=0.1, sig-dominated) and V8 (λsig=0.00025, action-dominated).

Architecture / Training

  • —Backbone: Qwen2.5-VL-3B-Instruct + LoRA (r=32, q/v_proj)
  • —Single AttentiveLatentHead A + ResNet action head
  • —Loss: 1.0·L1_action + 0.01·SIGReg(z_a, z_domain_current)
  • —100K steps, bs=64, lr=1e-4 cosine, 10 epochs

Results — epoch_10

SettingSR
AA0.72
BB0.72
C_A0.12
C_B0.24
AL_B0.0
BR_A0.0

Videos: 25 ep × 6 settings = 150 mp4.