CoolFace
Modelpublic

ubr-physical-ai/Cosmos3-Edge-INT4-AWQ

sourceHugging Faceotherupdated 16d agoView on Hugging Face
1likes188downloads
21 commits on main
0d7d69a16d ago

Add measured accuracy: INT4 indistinguishable from bf16 once k is re-fitted

filipemiguelmartins
347950b16d ago

Warn that thinking is off by default and breaks structured output (11/24 vs 24/24 parseable)

filipemiguelmartins
ed2205716d ago

Warn that thinking is off by default and breaks structured output (11/24 vs 24/24 parseable)

filipemiguelmartins
e6bcb3e16d ago

Card fixes: seven staging gaps not five; consistent GiB/GB units for the checkpoint

filipemiguelmartins
381941816d ago

Card fixes: seven staging gaps not five; consistent GiB/GB units for the checkpoint

filipemiguelmartins
4bb035b17d ago

First confirmed Jetson Orin Nano bring-up: 52.8 tok/s, 2.24 GiB engines, 2.7 GB peak

filipemiguelmartins
a4028ce17d ago

First confirmed Jetson Orin Nano bring-up: 52.8 tok/s, 2.24 GiB engines, 2.7 GB peak

filipemiguelmartins
e91a07a17d ago

Correct the reasoner subset: und_prefill is a policy component, not part of the VLM path

filipemiguelmartins
192bf1617d ago

Correct the reasoner subset: und_prefill is a policy component, not part of the VLM path

filipemiguelmartins
f61334517d ago

Fix the decoder MLP: 56 all-zero tensors were never bound (mlp.fc1/fc2 -> up_proj/down_proj)

filipemiguelmartins
c739ad517d ago

Publish the v2 checkpoint: projector excluded from quantisation

filipemiguelmartins
2cf9ea617d ago

Publish the v2 checkpoint: projector excluded from quantisation

filipemiguelmartins
b4bf9cd17d ago

Publish the v2 checkpoint: projector excluded from quantisation

filipemiguelmartins
48fd22b17d ago

Replace the broken ONNX export: real INT4 decoder (1.4 GB, 169 INT8 tensors), projector excluded

filipemiguelmartins
b98402617d ago

Correct the fit table and staging gaps: the 8.4 GB figure came from a broken export

filipemiguelmartins
11ab6ab17d ago

Generalise Jetson notes to Orin Nano 4 GB / 8 GB

filipemiguelmartins
c4a244517d ago

Generalise Jetson notes to Orin Nano 4 GB / 8 GB

filipemiguelmartins
cca2ffc17d ago

Generalise Jetson notes to Orin Nano 4 GB / 8 GB

filipemiguelmartins
1b494a217d ago

Add the Edge-LLM ONNX export (llm, visual, und_prefill, text_tokenizer); state the loading runtime, correct the 8-bit tag, add intended use and limitations

filipemiguelmartins
d27eda817d ago

INT4 AWQ quantisation of nvidia/Cosmos3-Edge for Jetson Orin (unvalidated)

filipemiguelmartins
6a605ad17d ago

initial commit

filipemiguelmartins