MicroLLM lab — tiny LLMs, Q4, in your browser
Objective checks (regex / exact tokens), not writing quality. A 135M model is allowed
to fail — that is the measurement. Pick models, then run. Estimate uses your
last tok/s if we have one.
Speed (tokens/s, sustained decode, suite wall) and accuracy (pass rate on objective tests) from runs
in this browser. Numbers stay on this machine. Charts use the latest suite per model.
Your Name or Handle
Device & Hardware Information
Write a benchmark in JavaScript
The editor is eval()’d in this origin, th...
Read more at stateofutopia.com