Spots

MicroLLM lab — tiny LLMs, Q4, in your browser

Objective checks (regex / exact tokens), not writing quality. A 135M model is allowed to fail — that is the measurement. Pick models, then run. Estimate uses your last tok/s if we have one.

Speed (tokens/s, sustained decode, suite wall) and accuracy

Speed (tokens/s, sustained decode, suite wall) and accuracy (pass rate on objective tests) from runs in this browser. Numbers stay on this machine. Charts use the latest suite per model. Verified Benchmark Certificate & Social Share

Generate and download a verifiable performance certificate with

Generate and download a verifiable performance certificate with your device hardware, peak and sustained tokens/second, and share your score. Write a benchmark in JavaScript The editor is eval()’d in this origin, then each check runs on the model’s decoded text.

News

MicroLLM lab — tiny LLMs, Q4, in your browser

Objective checks (regex / exact tokens), not writing quality.

@spots
Source: Hacker News
See more like this