智谱AI探索定制ASIC芯片,GLM-5.2使用量一周暴涨27倍

Rohan Paul · @rohanpaul_ai · X·2026-07-07 23:45·55天前
AI 导读

据The Information,智谱AI在DeepSeek之后也开始探索定制ASIC芯片,原因是其GLM-5.2使用量一周内暴涨27倍。定制ASIC虽牺牲灵活性,但可降低功耗和每token成本。Nvidia GPU适合通用场景,大规模推理需针对性设计。智谱尚未选定合作伙伴,项目或需两年以上。此前DeepSeek已自研推理芯片以摆脱对Nvidia和华为的依赖,但芯片工作尚处早期,大规模制造受美国限制先进代工厂和高带宽内存的制约。

Rohan Paul@rohanpaul_ai
50AI 编辑部评分,满分 100

智谱AI探索定制ASIC芯片,GLM-5.2使用量一周暴涨27倍

2026-07-07 23:45· 55天前
AI 导读

据The Information,智谱AI在DeepSeek之后也开始探索定制ASIC芯片,原因是其GLM-5.2使用量一周内暴涨27倍。定制ASIC虽牺牲灵活性,但可降低功耗和每token成本。Nvidia GPU适合通用场景,大规模推理需针对性设计。智谱尚未选定合作伙伴,项目或需两年以上。此前DeepSeek已自研推理芯片以摆脱对Nvidia和华为的依赖,但芯片工作尚处早期,大规模制造受美国限制先进代工厂和高带宽内存的制约。

Per The Information, Zhipu AI is also (after DeepSeek) exploring a custom ASIC after GLM-5.2 usage reportedly jumped 27x in one week.

A custom ASIC removes flexibility, but it can cut power draw and per-token cost.

Nvidia GPUs are strong general-purpose machines, but inference at scale has different economics. A fixed model can run better on silicon designed around its own repeated operations.

Zhipu has not chosen a partner, and the project may take more than 2 years.

The pattern is now bigger than one Chinese lab or one model launch. Chinese AI companies are trying to make software, hardware, and deployment less separable.

Rohan PaulDeepSeek is building an inference chip to cut dependence on Nvidia and Huawei in China’s $50B AI-chip market. DeepSeek’s chip work is still early, with outside ...