# OpenAI分析意外思维链评分对模型影响

- 来源：OpenAI (@OpenAI)
- 发布时间：2026-05-09 04:19
- AIHOT 分数：64
- AIHOT 链接：https://aihot.virxact.com/items/cmoxd7fk000tdsllhs7ybwwpf
- 原文链接：https://x.com/OpenAI/status/2052845764507062349

## AI 摘要

思维链监控器是防御AI智能体错位的关键层。为保持可监控性，我们在RL期间避免惩罚错位推理。

我们发现少量意外思维链评分影响了已发布模型，现分享相关分析。
https://alignment.openai.com/accidental-cot-grading/

## 正文

Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misaligned reasoning during RL.

We found a limited amount of accidental CoT grading which affected released models, and are sharing our analysis.
https://alignment.openai.com/accidental-cot-grading/
