mimile
Back to feed

SpaceXAI Grok 4.6 matches OpenAI GPT-5.6 Sol score at 60% lower cost

AI digest

This digest was compiled by AI from multiple sources — links to the originals are below.

SpaceXAI Grok 4.6 matches OpenAI GPT-5.6 Sol score at 60% lower cost

SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, tying OpenAI's GPT-5.6 Sol and trailing only Anthropic's Claude Opus 5 and Claude Fable 5. The model enters frontier competition at $2 per million input tokens and $6 per million output tokens, more than 60 percent below OpenAI and Anthropic flagship pricing. It also ranks second on the GDPval-AA v2 agentic benchmark, completing complex tasks in roughly 53 steps versus Claude Opus 5's 103.

Benchmark Performance

SpaceXAI's Grok 4.6 records a 61 on the Artificial Analysis Intelligence Index, the same score as OpenAI's GPT-5.6 Sol and one point behind Anthropic's Claude Fable 5. Only Claude Opus 5 (63) and Claude Fable 5 (62) rank higher. The result is a five-point gain over Grok 4.5, the previous version. Wccftech noted the model offers Fable 5-level performance at a fraction of the cost, a point echoed by The Decoder's index data.

Pricing and Availability

Grok 4.6 launches at $2 per million input tokens and $6 per million output tokens, unchanged from its predecessor. That is more than 60 percent below Claude Opus 5 ($5/$25) and GPT-5.6 Sol ($5/$30), according to The Decoder. Wccftech's headline described the discount as 80 percent, though its report did not detail the baseline. The model is available now via the API, Cursor, Grok Build, OpenRouter, Vercel, and Cloudflare. For the first week, x.ai offers double the usage quota in Grok Build and Cursor.

Agentic Workflows

On the GDPval-AA v2 benchmark, which measures real-world knowledge work on a computer, Grok 4.6 ranks second with an Elo score of 1,753, trailing only Claude Opus 5. The model completes complex tasks in about 53 steps, whereas Claude Opus 5 requires roughly 103. That efficiency positions SpaceXAI as a challenger in agentic multi-step workflows, where The Decoder says Grok 4.6 performs especially well.

What's Next

The immediate next test is sustained developer adoption across x.ai's partner platforms, including OpenRouter, Vercel, and Cloudflare. It remains unclear whether Anthropic or OpenAI will respond with price cuts or new releases to blunt Grok 4.6's cost advantage.

2 sources

SpaceXAI Grok 4.6 matches OpenAI GPT-5.6 Sol score at 60% lower cost