OpenBMB · @OpenBMB · X·2026-09-08 22:23·45分钟前
AI 导读

感谢测试并分享这些数据!🙌 “在本地运行一群这样的模型”——这正是我们的愿景。2B 参数,Q4 量化后 1.56GB,4090 上 200+ tok/s,同时性能仍能跟上 4B 模型。这就是本地 AI 智能体应有的样子。 期待看到你用它们构建的东西 🔥

OpenBMB@OpenBMB
34AI 编辑部评分,满分 100
2026-09-08 22:23· 45分钟前
AI 导读

感谢测试并分享这些数据!🙌 “在本地运行一群这样的模型”——这正是我们的愿景。2B 参数,Q4 量化后 1.56GB,4090 上 200+ tok/s,同时性能仍能跟上 4B 模型。这就是本地 AI 智能体应有的样子。 期待看到你用它们构建的东西 🔥

Thanks for testing it out and sharing these numbers! 🙌

"Run a swarm of these locally" — that's exactly the vision we had. 2B params, 1.56GB Q4, 200+ tok/s on a 4090, and still keeping up with 4B models. That's what local AI agents should look like.

Can't wait to see what you build with them 🔥 https://x.com/outsource_/status/2097005719983689998/video/1

来源:OpenBMB· x.com