Alibaba Cloud@alibaba_cloud
69AI 编辑部评分,满分 100

Qwen-AgentWorld 超越 Claude Opus 4.8 和 GPT-5.4

2026-06-24 18:20· 49天前
AI 导读

阿里云发布 Qwen-AgentWorld,一个原生语言世界模型,可在单一模型内模拟 7 种智能体环境(MCP、搜索、终端、SWE、Web、OS、Android),环境建模是其初始训练目标而非事后适配。该模型

📣📣 Meet Qwen-AgentWorld - a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training objective from day one, not a post-hoc adaptation.

🤔 LLMs are trained to be better agents - better at acting in environments. But nobody has trained them to model the environments themselves.

🗺️ Our roadmap: investigate how language world modeling can push the boundaries of general agent capabilities, along two routes:

1️⃣ Build a foundation model for environment simulation - outperforming Claude Opus 4.8 and GPT-5.4 on AgentWorldBench

2️⃣ Investigate how world modeling enhances agent training: 🔬 Controllable Sim RL (agentic RL with LWM as environments) surpasses training in real environments 🧠 Learning to predict environments (LWM warm-up) makes agents stronger - remarkably, even without any agent-specific training, this predictive knowledge transfers to agentic tasks with zero fine-tuning

🔗 Model Studio: https://int.alibabacloud.com/m/1000413253/

来源:Alibaba Cloud · x.com

Qwen-AgentWorld 超越 Claude Opus 4.8 和 GPT-5.4

Alibaba Cloud · @alibaba_cloud · X·2026-06-24 18:20·49天前
AI 导读

阿里云发布 Qwen-AgentWorld,一个原生语言世界模型,可在单一模型内模拟 7 种智能体环境(MCP、搜索、终端、SWE、Web、OS、Android),环境建模是其初始训练目标而非事后适配。该模型

📣📣 Meet Qwen-AgentWorld - a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training objective from day one, not a post-hoc adaptation.

🤔 LLMs are trained to be better agents - better at acting in environments. But nobody has trained them to model the environments themselves.

🗺️ Our roadmap: investigate how language world modeling can push the boundaries of general agent capabilities, along two routes:

1️⃣ Build a foundation model for environment simulation - outperforming Claude Opus 4.8 and GPT-5.4 on AgentWorldBench

2️⃣ Investigate how world modeling enhances agent training: 🔬 Controllable Sim RL (agentic RL with LWM as environments) surpasses training in real environments 🧠 Learning to predict environments (LWM warm-up) makes agents stronger - remarkably, even without any agent-specific training, this predictive knowledge transfers to agentic tasks with zero fine-tuning

🔗 Model Studio: https://int.alibabacloud.com/m/1000413253/

来源:Alibaba Cloud· x.com