SenseNova U1.5 Lite 开源:8B 原生统一多模态模型

Rohan Paul · @rohanpaul_ai · X·2026-08-21 01:10·5天前
AI 导读

SenseNova 发布 U1.5 Lite,一个 8B 参数的开源轻量级原生统一多模态模型,可同时完成视觉理解、生成与编辑。该模型通过 OPD 将任务专家能力蒸馏进单一模型,推理时无需路由或切换专家,并支持原生 4K 生成、复杂指令遵循及本地编辑。其在指令遵循和编辑保持上超越同尺寸模型,文本渲染与复杂布局上可媲美大型商业模型。

Rohan Paul@rohanpaul_ai
47AI 编辑部评分,满分 100

SenseNova U1.5 Lite 开源:8B 原生统一多模态模型

2026-08-21 01:10· 5天前
AI 导读

SenseNova 发布 U1.5 Lite,一个 8B 参数的开源轻量级原生统一多模态模型,可同时完成视觉理解、生成与编辑。该模型通过 OPD 将任务专家能力蒸馏进单一模型,推理时无需路由或切换专家,并支持原生 4K 生成、复杂指令遵循及本地编辑。其在指令遵循和编辑保持上超越同尺寸模型,文本渲染与复杂布局上可媲美大型商业模型。

SenseNova's full U1.5-Lite release is basically a transition from "how many things can one model do?" to "can it do them together without falling apart?"

An 8B-param, open-source, lightweight native unified multimodal model for visual understanding, generation, and editing.

It did not chase a bigger model with U1.5-Lite; it chased a model that could reliably combine more visual skills at the same time.

SenseNova U1.5 Lite treats specialization as a training problem, then gives users 1 model for complex prompts, native 4K, text rendering, and local edits.

And because editing is native to the unified model, the source image, edit target, and generated result stay inside the same model workflow.

SenseNova first trains task-specialized experts for text rendering and infographics, aesthetic quality, and image editing. OPD then transfers those capabilities into one lightweight unified model, so inference does not require a router, expert switching, or manual model selection.

The full release also applies task-oriented RL around instruction adherence, visual preference, and edit fidelity.

That maps directly to the visible improvements: stronger complex-prompt handling, better composition and text layouts, stable native 2K/4K high-resolution generation, and local edits that preserve identity, geometry, and untouched regions.

SenseTime𝗦𝗲𝗻𝘀𝗲𝗡𝗼𝘃𝗮 𝗨𝟭.𝟱 𝗟𝗶𝘁𝗲 — 𝗦𝗵𝗮𝗿𝗽𝗲𝗿. 𝗠𝗼𝗿𝗲 𝗖𝗼𝗻𝘁𝗿𝗼𝗹𝗹𝗮𝗯𝗹𝗲. 𝗠𝗼𝗿𝗲 𝗖𝗼𝘀𝘁-𝗘𝗳𝗳𝗶𝗰𝗶𝗲𝗻𝘁. An 𝗼𝗽𝗲𝗻-𝘀𝗼𝘂𝗿𝗰𝗲, lightwe...