跳到正文
原文
Hacker News:AI 热帖· Hacker News:AI 热帖·· 7 天前AI 评分49

MicroLLM Lab:在浏览器中试用 7 款微型 LLM

AI 导读

MicroLLM Lab 上线,可在浏览器内直接试用 7 款微型 LLM,并支持用 JavaScript 编写基准测试。测试以正则或精确 token 做客观校验,不评判写作质量,135M 模型允许失败,这本身就是测量结果。运行会记录 tokens/s 持续解码速度与客观测试通过率,数据只留在本机,图表按每个模型的最新测试套件展示。

正文

Next action:

Objective checks (regex / exact tokens), not writing quality. A 135M model is allowed to fail — that is the measurement. Pick models, then run. Estimate uses your last tok/s if we have one.

Next action:

Speed (tokens/s, sustained decode, suite wall) and accuracy (pass rate on objective tests) from runs in this browser. Numbers stay on this machine. Charts use the latest suite per model.

Next action:

Write a benchmark in JavaScript

The editor is eval()’d in this origin, then each check runs on the model’s decoded text.

Next action:

来源:Hacker News:AI 热帖 · stateofutopia.com