Tech

MicroLLM Lab brings small language model tests into the browser

The tool runs quantised models with WebGPU and reports browser-run speed and pass rates on objective checks, with results kept on the user’s machine.

Editorial persona
Owen Mercer
Markets and Finance Editor
Published
Draft
Source: Hacker News · View original source
Tech
No image available
Artificial intelligence

MicroLLM Lab offers a browser-based way to chat with and compare quantised small language models using WebGPU. Its description names PetitGPT and SmolLM2 among the models available; the headline advertises seven, but does not list them all.

The tool measures sustained decoding speed and pass rates on objective tests, including regex and exact-token checks. Those results reflect performance on the specified tests, rather than a broad assessment of writing quality.

MicroLLM Lab says measurements come from runs in the user’s browser and remain on the user’s machine. Charts display the latest test suite for each model. The tool may estimate performance using the user’s last recorded tokens-per-second result.

Users can download a performance certificate listing device hardware and peak and sustained tokens-per-second figures. The source provides no independent verification of the results.

The tool also says its editor is evaluated with `eval()` in the page’s origin before checks run against model output.

Continue reading

More from Tech

Read next: Google sets ChromeOS support through mid-2034 for enterprise and education
Read next: Discord plans Game Mode rollout to cut app resource use during gaming
Read next: Boox opens pre-orders for compact Picco e-reader at US$100