AI科学家应作为人机协作系统研究

Rohan Paul · @rohanpaul_ai · X·2026-08-21 00:19·4天前
AI 导读

一篇立场论文主张,科学智能体应作为“科学家+智能体”的人机协作系统来评估,而非孤立优化智能体本身。在10项科学任务中,GPT-5-mini几乎从不主动请求人类帮助,但专家能发现智能体遗漏的错误,而智能体加速执行——增益来自协作而非自主性。论文提出以人机团队是否比任何单独一方产出更好科学为基准。

Rohan Paul@rohanpaul_ai
35AI 编辑部评分,满分 100

AI科学家应作为人机协作系统研究

2026-08-21 00:19· 4天前
AI 导读

一篇立场论文主张,科学智能体应作为“科学家+智能体”的人机协作系统来评估,而非孤立优化智能体本身。在10项科学任务中,GPT-5-mini几乎从不主动请求人类帮助,但专家能发现智能体遗漏的错误,而智能体加速执行——增益来自协作而非自主性。论文提出以人机团队是否比任何单独一方产出更好科学为基准。

The race to build “AI Scientists” may be optimizing for the wrong unit: the agent alone.

This position paper argues that scientific agents should be studied as human-agent systems, where the thing you evaluate is the scientist + agent pair.

Most current systems still treat the human as a supervisor: set the goal, review a phase, approve the final artifact. Far fewer are built for continuous, fine-grained collaboration during the work itself.

The problem is that agents do not naturally know when they need human input. In 10 science tasks, GPT-5-mini almost never asked for help.

But that input matters: in the case studies, experts caught errors the agents missed, while the agents sped up execution. The gain came from collaboration, not autonomy.

The paper’s proposed benchmark is therefore different: does the human-agent team produce better science than either member alone, without collaboration cost overwhelming the gain?

– arxiv. org/abs/2608.14667

Title: "Position: AI Agents in Scientific Teams Should Be Studied as Human-Agent Systems"

来源:Rohan Paul· x.com