CoolFace
Modelpublic

0x8badbeef/molmo-audio-serving-blaster-think

sourceHugging Facecc-by-nc-sa-4.0updated 1mo agoView on Hugging Face
0likes3downloads
Model Card

Blaster-Think

Blaster-Think is a snapshot of the v35 hearing ship tip (checkpoint-116000) plus the Loop 2 typed-think decode contract. Weights are the same merged serving export as live Blaster at that tip. Thinking is not extra LoRA in this dump: the gateway prefills Answer: <think> on cued math and strips the span before TTS.

This repo is a separate Hub model so later GRPO / dual-LoRA work cannot overwrite the hearing ship:

SERVE_TIP on the live box stays outputs/molmo-audio-lora-diar-d-v35/checkpoint-116000. Think CE at 117500 was aborted (hearing dual-eval HEARING_NO_SHIP). Do not serve 117500.

What’s inside

Same files as the diar-d serving merge (model-*-of-*.safetensors, audio_modules.pt, tokenizer, remote-code sources), plus:

PathPurpose
serving/think_scaffold.pyLoop 2 P0/P1 cue, 192-token body, force-close, loop XOR
eval/think_eval_116000_gsm_192.jsonScaffolded GSM heldout at this tip
SNAPSHOT.jsonPin: tip id, parent, abort notes

Think eval (this tip, scaffolded GSM, n=20)

GateResult
Format (<think></think>)20/20
Exact8/20 (worded finals and a few real misses; not a ship gate for think)
Unearned correct0
Uncued leak0

Exact is a format contract + weak arithmetic, not a reason to invert training from base Molmo-7B-D.

Intended use

  • Restore this exact 116000 merge if GRPO or a second think LoRA goes wrong
  • Research on cued typed thinking without replacing live Blaster

License

[CC BY-NC-SA 4.0](https://creativecommons.org/licenses/by-nc-sa/4.0/). Private unless the owner changes visibility.

Marketing name: Blaster-Think. Technical Hub id: molmo-audio-serving-blaster-think.