CoolFace
Modelpublic

Solshine/gemma-4-e2b-nla-L23-ar-v0_1-paraphrase-invariant

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
0likes18downloads
13 commits on main
fd9faa73mo ago

Add v0.1 content-discrimination eval (routing vs within-domain content) + figure

Solshine
2237a603mo ago

Add v0.1 content-discrimination eval (routing vs within-domain content) + figure

Solshine
bd3babf3mo ago

v0.1 card: correct version + figures + verified facts

Solshine
bf9d6b53mo ago

v0.1 card: correct version + figures + verified facts

Solshine
da706453mo ago

v0.1 card: correct version + figures + verified facts

Solshine
11bb0004mo ago

Attribution: customized variation of the methodology + customizations section

Solshine
1e11b224mo ago

Soften polysemanticity language (evidence-against, ongoing) + drop 'training gap' framing; sync to GitHub

Solshine
af661a04mo ago

Reframe content-fidelity tone: verbalizer training gap, not content-blind (content present in activation, 60% probe, L17 lever); sync card to GitHub

Solshine
bdfdc614mo ago

Add Release rationale: why this SFT pair and not a GRPO checkpoint (self-contained section explaining the 2026-05-25 to 2026-05-29 Phase 4 GRPO trial outcome)

Solshine
5b8cd0e4mo ago

Add 2026-05-19 update: n=50 two-judge + 4-AR Δmse + polysemanticity caveat (§F72 Addenda 9-11)

Solshine
f9106a04mo ago

Update model card with calibrated post-Neuronpedia framing

Solshine
3e4f8f24mo ago

v0.1 paraphrase-invariance AR (model card + LoRA + linear_head + nla_meta)

Solshine
4829acb4mo ago

initial commit

Solshine