elvis · @omarsar0 · X·2026-09-11 00:00·44分钟前
AI 导读

Cognition 发布 SWE-2,称其在主要评测上与近期前沿模型相当,成本最多低 70%。SWE-2 在 FrontierCode 1.1 Main1 上得 50.0%,距 Fable 5.1 不到一分,同时便宜 64%;团队将 RL 扩展到多个万亿参数规模。

elvis@omarsar0
61AI 编辑部评分,满分 100
2026-09-11 00:00· 44分钟前
AI 导读

Cognition 发布 SWE-2,称其在主要评测上与近期前沿模型相当,成本最多低 70%。SWE-2 在 FrontierCode 1.1 Main1 上得 50.0%,距 Fable 5.1 不到一分,同时便宜 64%;团队将 RL 扩展到多个万亿参数规模。

Just amazed by how specialized models keep pushing the Pareto frontier.

Just watch that massive cost reduction.

SWE-2 achieves 50.0% on FrontierCode 1.1 Main1, within one point of Fable 5.1, while being 64% cheaper.

CognitionIntroducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL...