# Lambert反驳Thompson：中国实验室未用最强模型做RL蒸馏

- 来源：Nathan Lambert (@natolambert)
- 发布时间：2026-07-21 23:17
- AIHOT 分数：57
- AIHOT 链接：https://aihot.virxact.com/items/cmrusypah0bfybi9thf90zsp9
- 原文链接：https://x.com/natolambert/status/2079586505476165822

## AI 摘要

Yo @benthompson 抱歉，但中国实验室在强化学习阶段并没有使用 Fable / 最强模型作为教师，蒸馏不是这样运作的。

这样做不会带来那么大的提升（强化学习中的评分器本身就很混乱），而且你也负担不起那样使用 Fable。

## 正文

Yo @benthompson I'm sorry but the Chinese labs aren't using Fable / the strongest models as teachers during RL， that's not how distillation works.

It wouldn't give that big of a lift （graders during RL are messy） and you cant afford to use Fable like that.
