Grok 4.6 Takes #1 on CursorBench 3.2 β At 6Γ Lower Cost Than Fable 5
August 22, 2026
Grok 4.6 offers excellent value for money, and as we all know, xAI has produced a fantastic model with Cursor in a very short time. In many benchmarks, it performs on par with GPT-5.6, Opus 5, andβ¦
Grok 4.6 Takes #1 on CursorBench 3.2 β At 6Γ Lower Cost Than Fable 5
Grok 4.6 offers excellent value for money, and as we all know, xAI has produced a fantastic model with Cursor in a very short time. In many benchmarks, it performs on par with GPT-5.6, Opus 5, and even, in certain benchmarks, Fable 5.
However, we shouldn't forget that the full-size model is still to come. Grok 4.6 was just the 1.5T version.
"Grok 4.7 will be the 2.1T model released a few weeks later. This will be better than 4.6 in every way, except slightly slower to serve, albeit with even better token efficiency."
Kimi k3.1 is about to be released, GLM-5.3 (Flash) is currently demonstrating how good smaller models can be, and thus the pressure on OpenAI and Anthropic is increasing.
I'm very excited for Grok 4.7. Today I'll finally have more time to thoroughly test Grok Bot. I haven't had enough time so far. Grok 4.6 just took the #1 spot on CursorBench 3.2 β while delivering a massive efficiency advantage. β‘π»
β’ Grok 4.6 Extra High β 70.8% | $2.81/task β’ Fable 5 Max β 70.5% | $17.32/task β’ Opus 5 Max β 70.0% | $8.23/task β’ GPT-5.6 Sol Max β 67.2% | $5.69/task
Grok achieved the highest score while costing roughly 6Γ less than Fable 5 Max and nearly 3Γ less than Opus 5 Max per task.
For AI agents, raw intelligence is only part of the equation. The ability to maintain high performance across long coding tasks without burning massive amounts of compute could be a major advantage.
Grok's agentic coding efficiency is becoming seriously impressive. π
Source: CursorBench 3.2
Source: https://x.ai/news/grok-4-6