DogeDesigner@cb_doge
58AI 编辑部评分,满分 100
2026-08-13 01:42· 16小时前
AI 导读

Grok 4.6 在 ARI Bench 上夺得第一——该基准用于评测 AI 的递归式自我改进能力。🥇 它显著超越 Grok 4.5,同时排名也领先于 Claude Opus 5 和 GPT-5.6。

Grok 4.6 takes the #1 spot on ARI Bench - the benchmark for recursive AI improvement. 🥇

It significantly outperforms Grok 4.5 while also ranking ahead of Claude Opus 5 and GPT-5.6

来源:DogeDesigner · x.com

DogeDesigner · @cb_doge · X·2026-08-13 01:42·16小时前
AI 导读

Grok 4.6 在 ARI Bench 上夺得第一——该基准用于评测 AI 的递归式自我改进能力。🥇 它显著超越 Grok 4.5,同时排名也领先于 Claude Opus 5 和 GPT-5.6。

Grok 4.6 takes the #1 spot on ARI Bench - the benchmark for recursive AI improvement. 🥇

It significantly outperforms Grok 4.5 while also ranking ahead of Claude Opus 5 and GPT-5.6

来源:DogeDesigner· x.com