X:Rohan Paul (@rohanpaul_ai)· X:Rohan Paul (@rohanpaul_ai)·· 5 天前AI 评分43
webAI 3.66B 小模型 TwIL-LM3-Pro 登顶形式逻辑
AI 导读
美国 Austin 的 webAI 发布 3.66B 本地模型 TwIL-LM3-Pro,将 IBM Granite 的形式逻辑得分提升 28%,在形式逻辑上比 VibeThinker-3B 高约 35%、比 Qwen3.5-4B 高 24%、比 Liquid AI 的 LFM2.5-8B-A1B 高 47%。
正文
China's best small reasoning model (VibeThinker-3B) just got beaten by a 3.6B model from Austin.
webAI's TwIL-LM3-Pro, a 3.66B local model, lifts IBM Granite's formal-logic score by 28%
the recommended Q4 build is a 2.09GiB file that runs through llama.cpp on CPU or local GPU, so private data can stay on the device.
TwIL-LM3-Pro now leads every small model webAI compared on formal logic, scoring roughly 35% above Weibo's VibeThinker-3B, 24% above Qwen3.5-4B and 47% above Liquid AI's LFM2.5-8B-A1B.
来源:X:Rohan Paul (@rohanpaul_ai) · x.com