Rohan Paul · @rohanpaul_ai · X·2026-08-25 08:44·8小时前
AI 导读

Ox Alpha 三天内处理 11.6T tokens,是 OpenRouter 此前最大模型发布量的 2.6 倍,且有望今日突破 6T tokens。该模型拥有 1.05M-token 上下文窗口,专为持续智能体工作设计,单次会话可承载远超普通聊天的文本量。这种规模得益于编码智能体在单任务中通过长上下文、重试和重复工具调用大幅放大 token 消耗。

Rohan Paul@rohanpaul_ai
39AI 编辑部评分,满分 100
2026-08-25 08:44· 8小时前
AI 导读

Ox Alpha 三天内处理 11.6T tokens,是 OpenRouter 此前最大模型发布量的 2.6 倍,且有望今日突破 6T tokens。该模型拥有 1.05M-token 上下文窗口,专为持续智能体工作设计,单次会话可承载远超普通聊天的文本量。这种规模得益于编码智能体在单任务中通过长上下文、重试和重复工具调用大幅放大 token 消耗。

Ox Alpha processed 11.6T tokens in three days, 2.6x OpenRouter's previous biggest model launch.

This kind of scale is only possible now, because of coding agents, where long contexts, retries, and repeated tool calls can massively multiply token consumption inside one task.

Ox Alpha has a 1.05M-token context window and is intended for sustained agentic work, so each session can carry far more text than ordinary chat.

OpenRouterOx Alpha is on track to hit nearly 6 trillion tokens today. Try it now via ori in your coding agents: $ ori [your favorite harness] --model stealth/ox-alpha