Seedance 2.0 是一款全新的原生多模态音视频生成模型,于 2026 年 2 月初在中国正式发布。与之前的 Seedance 1.0 和 1.5 Pro 版本相比,Seedance 2.0 采用了统一、高效且大规模的多模态音视频联合生成架构。这使得它能够支持文本、图像、音频和视频四种输入模态,集成了业界迄今为止最全面的多模态内容参考与编辑功能套件之一。该模型在视频和音频生成的所有关键子维度上都实现了显著且全面的提升。在专家评估和公开用户测试中,该模型的表现已达到该领域的领先水平。Seedance 2.0 支持直接生成时长在 4 到 15 秒之间的音视频内容,原生输出分辨率为 480p 和 720p。对于作为参考的多模态输入,其当前开放平台最多支持 3 个视频片段、9 张图像和 3 个音频片段。此外,我们还提供了 Seedance 2.0 Fast 版本,这是 Seedance 2.0 的加速变体,旨在提升低延迟场景下的生成速度。Seedance 2.0 在其基础生成能力和多模态生成性能方面均实现了显著改进,为终端用户带来了增强的创作体验。
Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro, Seedance 2.0 adopts a unified, highly efficient, and large-scale architecture for multi-modal audio-video joint generation. This allows it to support four input modalities: text, image, audio, and video, by integrating one of the most comprehensive suites of multi-modal content reference and editing capabilities available in the industry to date. It delivers substantial, well-rounded improvements across all key sub-dimensions of video and audio generation. In both expert evaluations and public user tests, the model has demonstrated performance on par with the leading levels in the field. Seedance 2.0 supports direct generation of audio-video content with durations ranging from 4 to 15 seconds, with native output resolutions of 480p and 720p. For multi-modal inputs as reference, its current open platform supports up to 3 video clips, 9 images, and 3 audio clips. In addition, we provide Seedance 2.0 Fast version, an accelerated variant of Seedance 2.0 designed to boost generation speed for low-latency scenarios. Seedance 2.0 has delivered significant improvements to its foundational generation capabilities and multi-modal generation performance, bringing an enhanced creative experience for end users.