xAI 发布 Grok 4.7:编码与长时智能体能力提升,价格与 Grok 4.6 持平

Mark Kretschmann · @mark_k · X·2026-09-22 00:24·1小时前
AI 导读

Grok 4.7 发布,编码、长时运行的智能体和知识工作能力提升,价格与速度与 Grok 4.6 相同。主要改进包括更大的基础模型、更长强化学习训练、更好的自检与长上下文处理、更强的抗越狱能力,并已上架 Cursor、Grok Build 和 API。

Mark Kretschmann@mark_k
73AI 编辑部评分,满分 100

xAI 发布 Grok 4.7:编码与长时智能体能力提升,价格与 Grok 4.6 持平

2026-09-22 00:24· 1小时前
AI 导读

Grok 4.7 发布,编码、长时运行的智能体和知识工作能力提升,价格与速度与 Grok 4.6 相同。主要改进包括更大的基础模型、更长强化学习训练、更好的自检与长上下文处理、更强的抗越狱能力,并已上架 Cursor、Grok Build 和 API。

Grok 4.7 is here. @SpaceXAI’s latest release brings stronger coding, longer-running agents and better knowledge work, at the same price and speed as Grok 4.6.

The biggest improvements:

• A larger base model, trained with longer reinforcement learning on harder tasks that can take hours to complete. • Better self-checking and handling of long context. • Improved document and presentation creation. • Native understanding of the Grok Bot harness for better conversations and general knowledge work. • A rebuilt safeguard system with stronger jailbreak resistance and fewer unnecessary refusals for legitimate cybersecurity work.

The published benchmarks show meaningful gains over 4.6: • CursorBench 4.0: 40.4% → 46.3% • Terminal-Bench 4.0: 20.3% → 38.0% • EEBench: 53.0% → 64.0% • Harvey Legal Agent Benchmark: 15.8% → 19.6% • GDPval: 1,605 → 1,695 Elo

Against competitors, Grok 4.7 beats GPT-5.6 Sol on CursorBench and Terminal-Bench, and leads both Sol and Fable 5.1 on EEBench and Harvey. Fable still leads on CursorBench, Terminal-Bench and GDPval. Strong progress, with a particularly compelling price-performance story.

Pricing starts at $2 per million input tokens and $6 per million output tokens. A fast variant offers twice the output speed at twice the price.

Available now in Cursor, Grok Build and the API.