Grok 4.6登顶CursorBench,4.7将更强

Chubby♨️ · @kimmonismus · X·2026-08-22 17:34·1天前
AI 导读

Grok 4.6 以 70.8% 登顶 CursorBench 3.2,单任务成本仅 $2.81,比 Fable 5 Max(70.5%,$17.32)便宜约 6 倍,比 Opus 5 Max(70.0%,$8.23)便宜近 3 倍。xAI 称完整版 Grok 4.7 为 2.1T 参数模型,将在数周后发布,各方面优于 4.6(1.5T),仅服务速度略慢但 token 效率更高。

Chubby♨️@kimmonismus
42AI 编辑部评分,满分 100

Grok 4.6登顶CursorBench,4.7将更强

2026-08-22 17:34· 1天前
AI 导读

Grok 4.6 以 70.8% 登顶 CursorBench 3.2,单任务成本仅 $2.81,比 Fable 5 Max(70.5%,$17.32)便宜约 6 倍,比 Opus 5 Max(70.0%,$8.23)便宜近 3 倍。xAI 称完整版 Grok 4.7 为 2.1T 参数模型,将在数周后发布,各方面优于 4.6(1.5T),仅服务速度略慢但 token 效率更高。

Grok 4.6 offers excellent value for money, and as we all know, xAI has produced a fantastic model with Cursor in a very short time. In many benchmarks, it performs on par with GPT-5.6, Opus 5, and even, in certain benchmarks, Fable 5.

However, we shouldn't forget that the full-size model is still to come. Grok 4.6 was just the 1.5T version.

"Grok 4.7 will be the 2.1T model released a few weeks later. This will be better than 4.6 in every way, except slightly slower to serve, albeit with even better token efficiency."

Kimi k3.1 is about to be released, GLM-5.3 (Flash) is currently demonstrating how good smaller models can be, and thus the pressure on OpenAI and Anthropic is increasing.

I'm very excited for Grok 4.7. Today I'll finally have more time to thoroughly test Grok Bot. I haven't had enough time so far.

Tesla Owners Silicon ValleyBREAKING: Grok 4.6 just took the #1 spot on CursorBench 3.2 — while delivering a massive efficiency advantage. ⚡💻 • Grok 4.6 Extra High — 70.8% | $2.81/task • ...

来源:Chubby♨️· x.com