# Kimi K3 登顶 Arena 全栈编码测试

- 来源：Rohan Paul (@rohanpaul_ai)
- 发布时间：2026-07-29 05:20
- AIHOT 分数：48
- AIHOT 链接：https://aihot.virxact.com/items/cms5637hk00pdroehsisozm7t
- 原文链接：https://x.com/rohanpaul_ai/status/2082214719667249386

## AI 摘要

Kimi K3 (Max) 在 Arena 全栈编码测试中排名第一，超越 GPT-5.6 Sol (xHigh) 和 Claude Fable 5。该基准要求模型规划、编辑文件、运行命令、连接数据库与 API，并生成可部署的 Web 应用，由人工评估功能与可用性。

## 正文

Another win for Kimi K3 （Max）

Now it ranks 1st on Arena's fullstack coding test ahead of GPT-5.6 Sol （xHigh） and Claude Fable 5.

This benchmark goes beyond isolated code snippets by asking models to build working web applications across several connected steps.

Models must plan， edit files， run commands， connect databases， authentication and APIs， and produce a deployable application.

Human evaluators then compare the apps for functionality， usability and how closely they match the requested behaviour.

So to perform on this benchmark models need stronger coordination across frontend， backend and tool decisions during a complete build.

### 引用推文

> Arena.ai：Code Arena now measures fullstack capabilities! View overall rankings across AI models on full-stack web development tasks: multi-step reasoning, tool use, and ...
