Telegram NewsSubmit game
SpaceXAI's Grok 4.6 ties OpenAI's best model and costs far less Image Source: The-decoder

SpaceXAI’s Grok 4.6 ties OpenAI’s best model and costs far less

George Tsagkarakis Updated 2 min read
Contents 4 sections
We may include affiliate links in our content, meaning we could earn a commission—or receive blockchain-based assets—if you click a link and make a purchase or take a specific action. Additionally, we use generative AI to help draft and refine our posts for clarity and grammar. All content is fact-checked and reviewed by a human editor before publication.

Grok 4.6 finishes complex tasks in about 53 steps. Claude Opus 5 needs roughly 103. That gap says more about what SpaceXAI built than the headline benchmark score does.

Level with OpenAI, just under Anthropic

On the Artificial Analysis Intelligence Index, Grok 4.6 scores 61 points, tying OpenAI’s GPT-5.6 Sol. Only Anthropic’s Claude Opus 5 at 63 and Claude Fable 5 at 62 score higher.

That’s a five-point jump over Grok 4.5. It’s also a two-point deficit to the top of the index, which is thin enough that you shouldn’t pick a model on the index alone.

Agentic tasks are where the step count matters

Grok 4.6 does its best work on agentic tasks, the ones where a model runs a multi-step workflow on its own without a human nudging it back on track.

On GDPval-AA v2, a benchmark built to measure real knowledge work done on a computer, it comes in second with an Elo score of 1,753, behind Claude Opus 5.

The step count is the part worth watching. Roughly half the steps for a task means roughly half the tokens burned getting there, and anyone who has watched an agent loop rack up a bill knows that’s not a rounding error.

SpaceXAI's Grok 4.6 ties OpenAI's best model and costs far less
SpaceXAI's Grok 4.6 ties OpenAI's best model and costs far less 1 SpaceXAI's Grok 4.6 ties OpenAI's best model and costs far less

The price is the real argument

Pricing stays at $2/$6 per million tokens. Claude Opus 5 runs $5/$25 and GPT-5.6 Sol runs $5/$30, which puts Grok 4.6 more than 60 percent cheaper than both.

Output tokens are where that spread bites hardest, and output tokens are exactly what agentic workloads generate in volume. A model that’s a couple of points behind on a benchmark index but a quarter of the output cost is an easy call for most production work.

Where you can run it today

Grok 4.6 is available now through the API, Cursor, Grok Build and partners including OpenRouter, Vercel and Cloudflare.

For the first week, x.ai is doubling the usage quota in Grok Build and Cursor. If you’re running agents that bill by the step, that’s the window to check whether 53 steps holds up on your own workload rather than on a benchmark harness.

Share this article
George Tsagkarakis

George Tsagkarakis, known as Staycalm4now is a professional author in the crypto gaming industry since early 2018. He has experienced all the growth of Blockchain Gaming and helped multiple projects achieve their goals and established a player base. He is the co-founder of egamers.io and now the Founder and owner of CryptoGames.gg He is also the COO of MyStage, an…

More from AI Models & Releases

Subscribe
Notify of
0 Comments
Oldest
Newest Most Voted