CoolFace
Modelpublic

litert-community/Qwen3.5-4B

sourceHugging Faceapache-2.0updated 5d agoView on Hugging Face
3likes4.7kdownloads
15 commits on main
2345e165d ago

Chat template: accept the 0.18 content-parts form (string form unchanged); weights, tokenizer and executor metadata byte-identical

mlboydaisuke
f06d12722d ago

Card: definition line (LiteRT, 2026-09)

mlboydaisuke
747276d22d ago

manifest: Galaxy S26 GPU rows from the S4 gate sweep (Android recommendation where none existed)

mlboydaisuke
377952922d ago

Card: base_model_relation: quantized (list under the base model's Quantizations)

mlboydaisuke
e08f0a524d ago

Add measured Raspberry Pi 5 (CPU) section

mlboydaisuke
e5e71a81mo ago

manifest: Pixel 8a measured row

mlboydaisuke
769191f1mo ago

Mixed INT4: measured Pixel 8a (8 GB) CPU row

mlboydaisuke
55e4f111mo ago

manifest: add Mixed INT4 variant

mlboydaisuke
fd7b9921mo ago

Qwen3.5-4B Mixed INT4 (int4 b32 linears, int8 embed/lm_head, fp32act)

mlboydaisuke
07912d91mo ago

Add Mixed INT4 variant (LiteRT-LM #1658)

mlboydaisuke
d86bd171mo ago

Add litertlm_manifest.json — machine-readable deployment manifest (variant selection, backend recommendations, measured performance)

mlboydaisuke
099ec681mo ago

Card: state where GPU execution is verified, and scope the Android decode note

mlboydaisuke
623c0991mo ago

Qwen3.5-4B int8 (fp32act, 7-signature ladder)

mlboydaisuke
f0b26061mo ago

Qwen3.5-4B LiteRT-LM card

mlboydaisuke
70081561mo ago

initial commit

mlboydaisuke