The Decoder:AI News(RSS)
63AI 编辑部评分,满分 100

SpaceXAI's Grok 4.6 matches OpenAI's best model and undercuts it on price

2026-08-13 02:33· 4天前· Matthias Bastian

Grok 4.6 catches up to frontier models while costing far less. According to the Artificial Analysis Intelligence Index, SpaceXAI's new model scores 61 points, tying OpenAI's GPT-5.6 Sol. Only Anthropic's Claude Opus 5 (63) and Claude Fable 5 (62) score higher. That's a five-point jump over its predecessor, Grok 4.5.

Grok 4.6 ties the top tier in the AA Index, which rolls multiple benchmarks into a single score. | Image: Artificial Analysis

Grok 4.6 performs especially well on agentic tasks, where models independently carry out multi-step workflows. On the GDPval-AA v2 benchmark, which aims to measure real-world knowledge work on a computer, it ranks second with an Elo score of 1,753, trailing only Claude Opus 5. It completes complex tasks in about 53 steps. Claude Opus 5 needs roughly 103.

Pricing stays at $2/$6 per million tokens. That's more than 60 percent cheaper than Claude Opus 5 ($5/$25) and GPT-5.6 Sol ($5/$30). Grok 4.6 is available now through the APICursorGrok Build, and partners like OpenRouter, Vercel, and Cloudflare. For the first week, x.ai is offering double the usage quota in Grok Build and Cursor.

AA

xAI

来源:The Decoder:AI News(RSS) · the-decoder.com

同一事件 · 2

SpaceXAI's Grok 4.6 matches OpenAI's best model and undercuts it on price

The Decoder:AI News(RSS)·2026-08-13 02:33·4天前·Matthias Bastian
原文 · 保持原样,未翻译

Grok 4.6 catches up to frontier models while costing far less. According to the Artificial Analysis Intelligence Index, SpaceXAI's new model scores 61 points, tying OpenAI's GPT-5.6 Sol. Only Anthropic's Claude Opus 5 (63) and Claude Fable 5 (62) score higher. That's a five-point jump over its predecessor, Grok 4.5.

Grok 4.6 ties the top tier in the AA Index, which rolls multiple benchmarks into a single score. | Image: Artificial Analysis

Grok 4.6 performs especially well on agentic tasks, where models independently carry out multi-step workflows. On the GDPval-AA v2 benchmark, which aims to measure real-world knowledge work on a computer, it ranks second with an Elo score of 1,753, trailing only Claude Opus 5. It completes complex tasks in about 53 steps. Claude Opus 5 needs roughly 103.

Pricing stays at $2/$6 per million tokens. That's more than 60 percent cheaper than Claude Opus 5 ($5/$25) and GPT-5.6 Sol ($5/$30). Grok 4.6 is available now through the APICursorGrok Build, and partners like OpenRouter, Vercel, and Cloudflare. For the first week, x.ai is offering double the usage quota in Grok Build and Cursor.

AA

xAI

来源:The Decoder:AI News(RSS)· the-decoder.com

同一事件 · 2