CoolFace
Modelpublic

litert-community/LFM2.5-1.2B-JP

sourceHugging Faceotherupdated 6d agoView on Hugging Face
1likes1.7kdownloads
19 commits on main
83046e06d ago

Chat template: accept the 0.18 content-parts form (string form unchanged); weights, tokenizer and executor metadata byte-identical

mlboydaisuke
8bd3bb319d ago

Card: measured device / browser results (edge-compat, 2026-09)

mlboydaisuke
d10abec21d ago

Card: definition line (LiteRT, 2026-09)

mlboydaisuke
7fa5fd822d ago

manifest: same-device CPU control rows on the Galaxy S26 for the flagship GPU recommendation (S7 backfill 2026-09-05); the recommendation is rewritten from the measured pair (measured win, kept for prefill, or reversed to cpu)

mlboydaisuke
dd0b04e23d ago

Card: base_model_relation: quantized (list under the base model's Quantizations)

mlboydaisuke
3c44d0a24d ago

Add measured Raspberry Pi 5 (CPU) section

mlboydaisuke
bf387ed1mo ago

litertlm_manifest.json: add the _int8_gpu variant

mlboydaisuke
a1931131mo ago

Card: document the _int8_gpu file; correct the quantization sentence in Conversion notes

mlboydaisuke
66da81e1mo ago

Add LFM2.5-1.2B-JP_int8_gpu.litertlm: GPU-capable int8 (same weights as _int8, re-exported on the litert-torch 0.9.3 lineage)

mlboydaisuke
5284dd21mo ago

README: iOS Metal verified on iPhone 17 Pro — set maxNumTokens=1024 (the #3129 failure was a context-sizing issue, not a runtime bug)

mlboydaisuke
58542262mo ago

Document the _int4_gpu variant: backend, measured GPU speeds, conversion notes

mlboydaisuke
ea8e7452mo ago

Add GPU-capable int4 variant (litert-torch 0.9.3 export; full OpenCL delegation)

mlboydaisuke
17379752mo ago

Card: state the actual GPU-delegation blocker (INT64 ShortConv ops, 536/579 delegated)

mlboydaisuke
690199d2mo ago

Card: measured Performance table (M4 Max CPU/GPU + on-device where measured), accuracy note

mlboydaisuke
86e84e12mo ago

Card: note the litert-lm 0.15 ExecutorMetadata update (weights unchanged)

mlboydaisuke
00dc0ce2mo ago

Add ExecutorMetadata section required by litert-lm >= 0.15 (weights unchanged; still runs on 0.14)

mlboydaisuke
dc72bc42mo ago

Add ExecutorMetadata section required by litert-lm >= 0.15 (weights unchanged; still runs on 0.14)

mlboydaisuke
e57474e2mo ago

LFM2.5-1.2B-JP LiteRT-LM: int8 linears-only (GSM8K 65% vs bf16 63 = parity; conv-int8 costs this tune ~9pt) + int4 736MB; CPU backend; conv-state prefill fix baked in

mlboydaisuke
cf5b3362mo ago

initial commit

mlboydaisuke