Chubby♨️ · @kimmonismus · X·2026-07-06 19:51·56天前
AI 导读

腾讯今天发布了Hy3。在盲测中击败GLM-5.1,同时运行的活跃参数比竞争对手更少 295B MoE,21B活跃参数,256K上下文。而且这是基于普通GQA实现的。没有稀疏注意力,没有MLA。因此效率并非来自架构技巧,还有提升空间。 这比基准数字更让竞争对手担忧。 中国产出的高效模型真是疯狂。

Chubby♨️@kimmonismus
53AI 编辑部评分,满分 100
2026-07-06 19:51· 56天前
AI 导读

腾讯今天发布了Hy3。在盲测中击败GLM-5.1,同时运行的活跃参数比竞争对手更少 295B MoE,21B活跃参数,256K上下文。而且这是基于普通GQA实现的。没有稀疏注意力,没有MLA。因此效率并非来自架构技巧,还有提升空间。 这比基准数字更让竞争对手担忧。 中国产出的高效模型真是疯狂。

Tencent released Hy3 today. Beating GLM-5.1 in blind tests while running fewer active params than the models it's competing with

295B MoE, 21B active, 256K context. And it does this on plain GQA. No sparse attention, no MLA. So the efficiency isn't coming from architectural tricks yet, there's still headroom left.

That should worry the competition more than the benchmark numbers do.

It’s truly crazy what kind of efficient models are coming out of China.

来源:Chubby♨️· x.com