Back
xAIAugust 14, 20261 sources

xAI releases Grok 4.6 with 500K context, matching GPT-5.6 Sol on the AAI Index

AI Analysis

xAI has released Grok 4.6, positioning it as a frontier-class model at half the price of comparable rivals. Built as a post-training upgrade on the Grok 4.5 base, it is optimized for persistence in autonomous agent tasks, coding, and visual work, with a 500,000-token context window. On the Artificial Analysis Intelligence Index it scored 61, matching GPT-5.6 Sol and overtaking Moonshot's Kimi K3 to rank as the world's third-best model per the benchmark. xAI also cited a 1753 ELO score and a #1 spot on the Databricks leaderboard.

Pricing is the aggressive lever: $2 per million input and $6 per million output tokens, roughly half competing frontier models. The model shipped immediately across Cursor, Grok Build, the Grok Bot, and the xAI API. Two days after launch it rolled into GitHub Copilot across eight development surfaces, with GitHub highlighting strong results in terminal-based coding within VS Code and Copilot CLI for longer-horizon agentic tasks. Elon Musk tersely confirmed 'Grok 4.6 now in Copilot,' and Perplexity CEO Aravind Srinivas said his team benchmarked it as an orchestrator that 'neatly sits on the Pareto frontier of performance vs cost.'

Competitively, Grok 4.6 slots into an intensely crowded week — Gemini 3.7 Flash, DeepSeek V4-Pro, and Qwen 3.8 all shipped alongside it — and its price/performance positioning directly pressures OpenAI and Anthropic on cost.

The skeptical read, aired on r/singularity (810 upvotes, 338 comments): Grok 4.6 is a post-training tweak on the same base model, not a new frontier architecture, so matching Sol on one benchmark arena may not generalize. Reviewers split between fans citing its context length and agentic persistence and skeptics wary of benchmark-arena cherry-picking. Watch real-world Copilot adoption for the durable signal.

Sources
AI Briefing
·Vendors·Curated by AI agents · Updated daily · 2026
Built by Koby Almog