Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

335 points · 385 comments on HN · read original →

Points and comments are a snapshot, not live.

Grok 4.6 scores 61 on the AI Intelligence Index, matching GPT-5.6 Sol at lower cost.

SpaceXAI's Grok 4.6 joins the frontier with an Intelligence Index score of 61, behind only Claude Opus 5 (63) and Claude Fable 5 (62). Headline pricing is unchanged from Grok 4.5 at $2/$6 per 1M input/output tokens, 60%+ below Claude Opus 5 and GPT-5.6 Sol. On agentic benchmarks, it achieves a GDPval-AA v2 Elo of 1753 (second only to Claude Opus 5) and scores 50.7% on τ³-Banking. It averages ~53 turns and ~0.5B input tokens on long-horizon tasks vs. ~103 turns and ~2.0B for Claude Opus 5.

What commenters are saying

Commenters are split on whether Grok 4.6 is truly competitive with Claude Opus 5 and GPT-5.6 Sol. Several note Cursor's subsidized subscription offers better value for Grok models than OpenAI or Anthropic plans. Some users report Grok 4.6 matches Claude in coding quality while being 3x faster, but others find it behind in edge-case testing. A minority speculates about weight distillation from Anthropic models hosted on SpaceXAI infrastructure.