非营利组织 Guidelight 首份评估:Anthropic、OpenAI 等 AI 实验室均未完全落实内部系统管控

The Decoder:AI News(RSS)·2026-08-19 21:16·4天前·Maximilian Schreiner
AI 导读

非营利组织 Guidelight 的首份评估显示,Anthropic、OpenAI、Google、xAI 和 Meta 均未完全落实针对内部 AI 系统的基本管控措施。评估基于系统卡、安全报告等公开资料,检查了日志记录、审查机制、紧急关停等六项实践。Anthropic 和 OpenAI 获 C+ 领先,Google 为 D+,xAI(D−)与 Meta(F)垫底。

The Decoder:AI News(RSS)
56AI 编辑部评分,满分 100

非营利组织 Guidelight 首份评估:Anthropic、OpenAI 等 AI 实验室均未完全落实内部系统管控

2026-08-19 21:16· 4天前· Maximilian Schreiner
AI 导读

非营利组织 Guidelight 的首份评估显示,Anthropic、OpenAI、Google、xAI 和 Meta 均未完全落实针对内部 AI 系统的基本管控措施。评估基于系统卡、安全报告等公开资料,检查了日志记录、审查机制、紧急关停等六项实践。Anthropic 和 OpenAI 获 C+ 领先,Google 为 D+,xAI(D−)与 Meta(F)垫底。

No AI company fully applies basic control measures to its own internal AI systems. That's the takeaway from the first assessment by the nonprofit Guidelight. The group looked at Anthropic, OpenAI, Google, xAI, and Meta, drawing only on public sources like system cards, safety reports, and blog posts.

Guidelight checked six basic practices. These include logging internal AI activity, gating risky actions through a review mechanism, emergency shutdowns known as "circuit breaking," and plans to contain misaligned models. Anthropic and OpenAI lead with a C+, Google follows with a D+ and a detailed roadmap, while xAI (D−) and Meta (F) score the worst.

Kein Unternehmen erreicht die vorgeschlagenen Sicherheit-Standards von Guidelight. | Bild: Guidelight
No company meets Guidelight's proposed safety standards. | Image: Guidelight

The companies do best at spotting misbehavior. They do worst at prevention and containment. Guidelight is an independent nonprofit founded by former OpenAI safety leads Page Hedley and Steven Adler.

Guidelight

来源:The Decoder:AI News(RSS)· the-decoder.com