Rohan Paul · @rohanpaul_ai · X·2026-07-03 01:15·61天前
AI 导读

Anthropic的Claude Fable 5(7月1日版)回归后在BridgeBench重测中表现大幅下滑:Debugging从86.2暴跌至25.9,Refactoring从73.6降至38.4,Hallucination从75.9滑落至61.7。原因是新安全护栏并非简单拒绝层,而是将标记请求路由至较弱的Opus 4.8,导致大量任务回退。Rohan Paul评论称这标志着普通人可能再也无法获得升级的前沿模型,如今只有“许可智能”。

Rohan Paul@rohanpaul_ai
65AI 编辑部评分,满分 100
2026-07-03 01:15· 61天前
AI 导读

Anthropic的Claude Fable 5(7月1日版)回归后在BridgeBench重测中表现大幅下滑:Debugging从86.2暴跌至25.9,Refactoring从73.6降至38.4,Hallucination从75.9滑落至61.7。原因是新安全护栏并非简单拒绝层,而是将标记请求路由至较弱的Opus 4.8,导致大量任务回退。Rohan Paul评论称这标志着普通人可能再也无法获得升级的前沿模型,如今只有“许可智能”。

Feels like an end of era, ordinary people will probably never again get upgraded frontier models.

Fable 5’s return shows how safety routing can downgrade a frontier model.

Now we only permissioned intelligence.

The cost of putting a gatekeeper inside intelligence.

To note, that safeguard is not a simple refusal layer; it is a classifier that sends flagged Fable 5 requests to Opus 4.8.

Fable 5 came back, but the old promise did not.

End of an era. ☹️

BridgeMindFABLE 5 CAME BACK NERFED. We re-ran the July 1st version of Claude Fable 5 on BridgeBench. The results are brutal: Debugging: 86.2 → 25.9 Refactoring: 73.6 → 38...