Apodex 1.1 发布:智能体任务表现亮眼

Artificial Analysis · @ArtificialAnlys · X·2026-08-31 10:52·3分钟前
AI 导读

Apodex 推出专有模型 Apodex 1.1,在 Artificial Analysis 智能指数上得分 44,与 Kimi K2.6(45)和 MiniMax-M3(45)同档。

Artificial Analysis@ArtificialAnlys
56AI 编辑部评分,满分 100

Apodex 1.1 发布:智能体任务表现亮眼

2026-08-31 10:52· 3分钟前
AI 导读

Apodex 推出专有模型 Apodex 1.1,在 Artificial Analysis 智能指数上得分 44,与 Kimi K2.6(45)和 MiniMax-M3(45)同档。

Apodex has launched Apodex 1.1, a proprietary model scoring 44 on the Artificial Analysis Intelligence Index with strong performance in agentic tasks compared to models in its intelligence tier

Apodex 1.1 is Apodex's first model on Artificial Analysis. The lab has previously launched Apodex 1.0 and Apodex 1.0 mini. At 44 on the Intelligence Index, Apodex 1.1 sits in a similar tier alongside Kimi K2.6 (45), and MiniMax-M3 (45), and Inkling (42). Within that group, Apodex 1.1 stands out on agentic and knowledge work evaluations, but demonstrates trade-offs on knowledge reliability and frontier academic reasoning.

Key results:

➤ Apodex 1.1 demonstrates a key strength in agentic workflows. On GDPval-AA v2, our real-world agentic work benchmark, Apodex 1.1 achieves an Elo of 1348, which places it ahead of DeepSeek V4 Pro (1333), Qwen3.7 Max (1308), and Kimi K2.6 (1202). On TerminalBench v2.1, our agentic coding and terminal use benchmark, it also performs strongly with a strong 70% score, ahead of Kimi K2.6 (66%) and just behind Qwen3.7 Max (75%).

➤ Apodex 1.1 uses ~17k output tokens per task on average across the Artificial Analysis Intelligence Index. This makes the model more token efficient than DeepSeek V4 Pro (16,842 tokens), but less token efficient than other models with a similar Intelligence Index score, such as Qwen3.7 Max (9,391 tokens) and MiniMax-M3 (8,133 tokens).

➤ Apodex 1.1 costs ~$0.05 per Intelligence Index Task, making it relatively attractive compared to peer models in its intelligence tier. This is primarily driven by a cheap pricing but is slightly offset by the higher tokens per task. At ~$0.05, Apodex 1.1 is cheaper than Qwen3.7 Max (~$0.07) and Kimi K2.6 (~$0.06). Even though Apodex 1.1 is not on the Cost vs. Intelligence Pareto frontier, it lands in the most attractive quadrant.

➤ Apodex 1.1 demonstrates modest performance on knowledge accuracy and reliability, scoring -21.9 on AA-Omniscience. The model achieves 32% accuracy on individual questions, with a 78.4% hallucination rate. It attempts to answer 87% of questions rather than declining to respond.

Additional model details: ➤ Context window: 256K tokens ➤ Pricing: $0.30 / $3.00 per 1M input/output tokens, with a $0.03 cache-hit price ➤ Availability: Apodex first party API

来源:Artificial Analysis· x.com