Interesting. Grok 4.7 is performing fairly well for a smallish model.
We ran fresh Next.js evals. The tally: ① Opus 5.5 [𝟿𝟽%] ② GPT 6 Sol [𝟿𝟽%] ③ Fable 5.1 [𝟿𝟽%] ④ Grok 4.7 [𝟿𝟺%] Notably, Grok is 2x-7x cheaper
有意思。Grok 4.7 作为一个小型模型表现相当不错。 (引用 @rauchg):我们跑了最新的 Next.js 评测。结果:① Opus 5.5 [97%] ② GPT 6 Sol [97%] ③ Fable 5.1 [97%] ④ Grok 4.7 [94%] 值得注意的是,Grok 便宜 2x-7x。
有意思。Grok 4.7 作为一个小型模型表现相当不错。 (引用 @rauchg):我们跑了最新的 Next.js 评测。结果:① Opus 5.5 [97%] ② GPT 6 Sol [97%] ③ Fable 5.1 [97%] ④ Grok 4.7 [94%] 值得注意的是,Grok 便宜 2x-7x。
Interesting. Grok 4.7 is performing fairly well for a smallish model.
来源:Elon Musk· x.com