Anthropic 内部运行未发布模型"Model 2",性能超越所有公开版 Claude

The Decoder:AI News(RSS)·2026-08-20 18:04·3天前·Maximilian Schreiner
AI 导读

Anthropic 在内部运行一款未发布的 AI 模型“Model 2”,其综合性能略强于 Claude Mythos 5,但在某些领域较弱。据该公司 2026 年 8 月的风险报告,该模型在内部能力指数 AECI 上比 Mythos 5 高约 1.5 点,提升幅度小于从 Mythos Preview 到 Mythos 5 的跃升。该模型主要用于内部编码、数据生成及研究与工程,目前无外部发布计划。

The Decoder:AI News(RSS)
57AI 编辑部评分,满分 100

Anthropic 内部运行未发布模型"Model 2",性能超越所有公开版 Claude

2026-08-20 18:04· 3天前· Maximilian Schreiner
AI 导读

Anthropic 在内部运行一款未发布的 AI 模型“Model 2”,其综合性能略强于 Claude Mythos 5,但在某些领域较弱。据该公司 2026 年 8 月的风险报告,该模型在内部能力指数 AECI 上比 Mythos 5 高约 1.5 点,提升幅度小于从 Mythos Preview 到 Mythos 5 的跃升。该模型主要用于内部编码、数据生成及研究与工程,目前无外部发布计划。

Anthropic is running an unreleased AI model internally that outperforms every publicly available version of Claude. That's according to the company's Risk Report from August 2026.

The report calls the model "Model 2" and places it in the Mythos class. Anthropic says it's slightly stronger overall than Claude Mythos 5, but weaker in some areas. It doesn't show a big capability jump like the one from Opus 4.6 to Mythos. On the company's internal capability index, AECI, it sits about 1.5 points above Mythos 5. That gain is smaller than the jump from Mythos Preview to Mythos 5.

Anthropic's ECI is a collection of internal benchmarks. "Model 2" is said to land 1.5 points above Mythos 5, but it isn't plotted here. | Image: Anthropic

Internally, the company leans on the model heavily for coding, data generation, and research and engineering, sometimes through agents that run continuously. Claude now writes most of the code in Anthropic's production systems. Model 2 went through an internal review before deployment, but wasn't tested as thoroughly as Mythos 5. Anthropic found no new or more worrying misalignments in the process. The company rates the overall risk from misalignment as "low." There are no plans to release the model externally right now.

来源:The Decoder:AI News(RSS)· the-decoder.com