CoolFace
Datasetpublic

NagaYu/molt-benchmark-results

Molt · elastic on-device inference measurements Everything measured while building Molt, a runtime that moves a running generation onto a smaller model between two tokens, carrying the KV cache across, so an on-device LLM under memory pressure is neither reclaimed by the OS nor restarted from the prompt. Published so the claims can be checked rather than taken on trust. The figures in the repo README and the results page are generated from these files; nothing is transcribed by… See the full description on the dataset page: https://huggingface.co/datasets/NagaYu/molt-benchmark-results.

sourceHugging Facemitupdated 21d agoView on Hugging Face
0likes54downloads
2 commits on main
2b0efd321d ago

Measurements: conditions, cost sweep, per-token series, projector residuals, raw logs

NagaYu
404df5f21d ago

initial commit

NagaYu