论文揭示智能体技能的真实作用机制

elvis · @omarsar0 · X·2026-08-17 23:39·19天前
AI 导读

一篇论文通过8,135条标准化试验记录揭示,智能体技能起作用时65.7%靠程序性锚定,仅4.5%靠显式知识注入,即技能主要稳定执行而非补充知识。技能库从5个增至100个时,实际使用精确率从29.6%降至3.3%;技能仍比Workflow Memory高6.06分,但在脆弱假设或不兼容情境下会失效。

elvis@omarsar0
42AI 编辑部评分,满分 100

论文揭示智能体技能的真实作用机制

2026-08-17 23:39· 19天前
AI 导读

一篇论文通过8,135条标准化试验记录揭示,智能体技能起作用时65.7%靠程序性锚定,仅4.5%靠显式知识注入,即技能主要稳定执行而非补充知识。技能库从5个增至100个时,实际使用精确率从29.6%降至3.3%;技能仍比Workflow Memory高6.06分,但在脆弱假设或不兼容情境下会失效。

Interesting paper demystifying agent skills.

If you maintain skills for your agent, this one is worth your time.

(bookmark it)

Skills are usually assumed to inject knowledge the model lacks. However, this paper finds something interesting.

Across 8,135 normalized trial records, procedural anchoring accounts for 65.7% of cases where a skill helps, and explicit knowledge injection accounts for 4.5%. Skills stabilize execution rather than supply facts.

As the pool grows from 5 to 100 skills, actual-use precision falls from 29.6% to 3.3%.

Skills still beat Workflow Memory by 6.06 points in matched comparisons, and they break under brittle assumptions, incompatible contexts, or insufficient adaptation.

Paper: https://arxiv.org/abs/2608.14036

Track more trending AI papers in our academy: https://academy.dair.ai/

来源:elvis· x.com