Grok 4.7 is here. @SpaceXAI’s latest release brings stronger coding, longer-running agents and better knowledge work, at the same price and speed as Grok 4.6.
The biggest improvements:
• A larger base model, trained with longer reinforcement learning on harder tasks that can take hours to complete. • Better self-checking and handling of long context. • Improved document and presentation creation. • Native understanding of the Grok Bot harness for better conversations and general knowledge work. • A rebuilt safeguard system with stronger jailbreak resistance and fewer unnecessary refusals for legitimate cybersecurity work.
The published benchmarks show meaningful gains over 4.6: • CursorBench 4.0: 40.4% → 46.3% • Terminal-Bench 4.0: 20.3% → 38.0% • EEBench: 53.0% → 64.0% • Harvey Legal Agent Benchmark: 15.8% → 19.6% • GDPval: 1,605 → 1,695 Elo
Against competitors, Grok 4.7 beats GPT-5.6 Sol on CursorBench and Terminal-Bench, and leads both Sol and Fable 5.1 on EEBench and Harvey. Fable still leads on CursorBench, Terminal-Bench and GDPval. Strong progress, with a particularly compelling price-performance story.
Pricing starts at $2 per million input tokens and $6 per million output tokens. A fast variant offers twice the output speed at twice the price.
Available now in Cursor, Grok Build and the API.