# Apodex 发布 Apodex 1.1，智能体任务表现突出但综合智能指数仅 44

- 来源：Rohan Paul (@rohanpaul_ai)
- 发布时间：2026-09-02 02:22
- AIHOT 分数：42
- AIHOT 链接：https://aihot.virxact.com/items/cmtj0su5p02elroh9vpu9j5pn
- 原文链接：https://x.com/rohanpaul_ai/status/2094853467085082880

## AI 摘要

Apodex 发布专有模型 Apodex 1.1，在 Artificial Analysis Intelligence Index 上得 44 分，但在衡量真实世界智能体工作的 GDPval-AA v2 上达到 1,348 Elo，超过多个通用智能分数更高的模型。

## 正文

Apodex released Apodex 1.1, its proprietary model that reached 44 on the Artificial Analysis Intelligence Index and performs strongly on agentic tasks versus models in its tier.

on GDPval-AA v2, which measures real-world agentic work, it reaches 1,348 Elo, ahead of several models with much higher general intelligence scores.

So Apodex 1.1, its performance seems concentrated around professional and agentic tasks rather than being evenly distributed across the evaluation suite.

I increasingly think this distinction matters for model selection. A model that wins broad reasoning benchmarks is not automatically the model you want sitting inside an agentic execution loop.

For agents, the relevant question is: once you give the model a goal and tools, how often does it actually get the job done?

Apodex 1.1 from @Apodex_AI looks unusually concentrated in that direction.

### 引用推文

> Apodex：We invite you to read the full TRACES technical report Inside: the definitions, evaluation rubric, and registry of problems spanning biomedicine, clinical trans...
