Grok 4.7 is our most capable model for coding and knowledge work. It works longer on difficult tasks, checks its own work more carefully, and comes with our best-calibrated safeguards to date. Served at the same price and speed as Grok 4.6 , it is highly competitive in its class. A scatter and line chart comparing Fable 5.1, Opus 5, Grok 4.7, GPT-5.6 Sol, and Sonnet 5 scores against average cost per task. 55% CursorBench 4.0 score 50% 45% 40% 35% 30% 25% 20% $18 $15 $12 $9 $6 $3 $0 Average cost per task Fable 5.1 Opus 5 GPT-5.6 Sol Sonnet 5 Grok 4.7 On CursorBench 4.0, which stresses longer-running coding tasks, Grok 4.7 is at the frontier in price-performance. Model Improvements Grok 4.7 uses a new, larger base model compared to Grok 4.6 . It was trained with a longer reinforcement learning run on a harder mix of tasks, weighted toward problems that take many hours to complete. The model is better at verifying its own work and managing longer context.…