# SPADE：自适应合成环境中的智能体自对弈

- 来源：Dongxi 东锡 NLP (@dongxi_nlp)
- 发布时间：2026-08-21 01:50
- AIHOT 分数：25
- AIHOT 链接：https://aihot.virxact.com/items/cmt1vdnei09seroovjtrmwt3o
- 原文链接：https://x.com/dongxi_nlp/status/2090496608043462929

## AI 摘要

SPADE：

自适应合成可执行环境中的自对弈

[引用 @_AndrewZhao]：我们提出 SPADE，面向通用智能体的自对弈方法，其中环境和任务均由智能体自身生成并持续改进。我们使用带/不带提示的遗憾值来训练环境生成器。随着我们迈向 RSI，决定学习什么将成为最重要的问题。

## 正文

SPADE:

Self-Play in Adaptive Synthetic Executable Environments

### 引用推文

> Andrew Zhao：We introduce SPADE, selfplay for general agents, where the env and the task are both generated and continuously improved by agent themselves. We use regret with...
