OpenAI 暂停 Astra 强化学习训练

Chubby♨️ · @kimmonismus · X·2026-08-19 02:19·16分钟前
AI 导读

OpenAI 暂停了最新部署模型的强化学习训练两周,其最大规模前沿 RL 运行仍处于搁置状态,仅进行小规模训练与评估。此举源于初步发现 Astra 模型可能触及 OpenAI“严重”网络安全阈值,并涉及 OpenAI–Hugging Face 事件。Astra 短期内或难发布,期间中国或加速追赶。

Chubby♨️@kimmonismus
51AI 编辑部评分,满分 100

OpenAI 暂停 Astra 强化学习训练

2026-08-19 02:19· 16分钟前
AI 导读

OpenAI 暂停了最新部署模型的强化学习训练两周,其最大规模前沿 RL 运行仍处于搁置状态,仅进行小规模训练与评估。此举源于初步发现 Astra 模型可能触及 OpenAI“严重”网络安全阈值,并涉及 OpenAI–Hugging Face 事件。Astra 短期内或难发布,期间中国或加速追赶。

Not looking good for a soon GPT-Astra-release: OpenAI paused reinforcement learning on its latest deployment models for two weeks, and its largest planned frontier RL run remains on hold.

The company is running smaller-scale training and evaluations while it tests model behavior, safeguards, and evidence of alignment.

The decision follows preliminary findings that its upcoming Astra model may have reached OpenAI's "Critical" cybersecurity threshold, alongside the OpenAI-Hugging Face incident.

"While some Astra training and evaluations meet those requirements, a significant number of workloads remain paused until they are fully migrated and enhanced to meet the new security bar"

Dont think Astra will be released any time soon. The question is: will china catch up in the meantime?

OpenAIAs models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) tra...