
MicroLLM lab — tiny LLMs, Q4, in your browser
Objective checks (regex / exact tokens), not writing quality. A 135M model is allowed to fail — that is the measurement. Pick models, then run. Estimate uses your last tok/s if we have one.
Objective checks (regex / exact tokens), not writing quality.