MicroLLM Lab:在浏览器中试用 7 款微型 LLM
MicroLLM Lab 上线,可在浏览器内直接试用 7 款微型 LLM,并支持用 JavaScript 编写基准测试。测试以正则或精确 token 做客观校验,不评判写作质量,135M 模型允许失败,这本身就是测量结果。运行会记录 tokens/s 持续解码速度与客观测试通过率,数据只留在本机,图表按每个模型的最新测试套件展示。
Next action:
Objective checks (regex / exact tokens), not writing quality. A 135M model is allowed to fail — that is the measurement. Pick models, then run. Estimate uses your last tok/s if we have one.
Next action:
Speed (tokens/s, sustained decode, suite wall) and accuracy (pass rate on objective tests) from runs in this browser. Numbers stay on this machine. Charts use the latest suite per model.
Next action:
Write a benchmark in JavaScript
The editor is eval()’d in this origin, then each check runs on the
model’s decoded text.
Next action:
来源:Hacker News:AI 热帖 · stateofutopia.com