中国开源模型成对齐研究主流基底

Dongxi 东锡 NLP · @dongxi_nlp · X·2026-08-25 01:01·1天前
AI 导读

对221个MATS对齐与安全项目的分析显示,2026年72%使用命名模型的项目采用中国开源权重模型,高于美国模型的54%;Qwen占66%,领先Llama(44%)和DeepSeek(32%)。闭源模型仍出现在90%项目中,多作为裁判、监控或前沿基线。中国开源模型正成为对齐研究的实验基底,而美国闭源模型仍是关键评估工具。

Dongxi 东锡 NLP@dongxi_nlp
61AI 编辑部评分,满分 100

中国开源模型成对齐研究主流基底

2026-08-25 01:01· 1天前
AI 导读

对221个MATS对齐与安全项目的分析显示,2026年72%使用命名模型的项目采用中国开源权重模型,高于美国模型的54%;Qwen占66%,领先Llama(44%)和DeepSeek(32%)。闭源模型仍出现在90%项目中,多作为裁判、监控或前沿基线。中国开源模型正成为对齐研究的实验基底,而美国闭源模型仍是关键评估工具。

I ran a similar analysis on 221 MATS alignment and safety projects.

In 2026, 72% of projects using named models use a Chinese open-weight family, versus 54% using an American one. Qwen appears in 66%, ahead of Llama at 44% and DeepSeek at 32%.

The interesting difference is that closed models still appear in 90% of these projects, often as judges, monitors, or frontier baselines.

Chinese open models are becoming the experimental substrate of alignment research, while closed American models remain key evaluation tools.

Nathan LambertOver the weekend I had Codex parse 500K arXiv AI/ML papers since ChatGPT to understand which open models are used for research. In 2024, ~30% of papers mentione...

来源:Dongxi 东锡 NLP· x.com