Seedance 2.0:面向世界复杂性的视频生成新进展

Seedance 2.0: Advancing Video Generation for World Complexity

高宇 Yu Gao · · 2026-04-15 · arXiv:2604.14148 ↗

打开互动全文版(逐段中英对照 + 图/公式 + 论文问答)→

摘要 · Abstract

Seedance 2.0 是一款新的原生多模态音视频生成模型,于 2026 年 2 月初在中国正式发布。与之前的 Seedance 1.0 和 1.5 Pro 相比,Seedance 2.0 采用了统一、高效且大规模的多模态音视频联合生成架构。这使得它能够支持文本、图像、音频和视频四种输入模态,并集成了迄今为止业界最全面的多模态内容参考与编辑能力之一。它在视频和音频生成的各个关键子维度上都带来了全面而显著的提升。在专家评测和公开用户测试中,该模型均表现出与该领域领先水平相当的性能。Seedance 2.0 支持直接生成时长 4 至 15 秒的音视频内容,原生输出分辨率包括 480p 和 720p。对于作为参考的多模态输入,其当前开放平台最多支持 3 个视频片段、9 张图像和 3 个音频片段。此外,我们还提供了 Seedance 2.0 Fast 版本,这是 Seedance 2.0 的加速变体,旨在提升低延迟场景下的生成速度。Seedance 2.0 在其基础生成能力和多模态生成性能方面均实现了显著改进,为最终用户带来了更出色的创作体验。

Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro, Seedance 2.0 adopts a unified, highly efficient, and large-scale architecture for multi-modal audio-video joint generation. This allows it to support four input modalities: text, image, audio, and video, by integrating one of the most comprehensive suites of multi-modal content reference and editing capabilities available in the industry to date. It delivers substantial, well-rounded improvements across all key sub-dimensions of video and audio generation. In both expert evaluations and public user tests, the model has demonstrated performance on par with the leading levels in the field. Seedance 2.0 supports direct generation of audio-video content with durations ranging from 4 to 15 seconds, with native output resolutions of 480p and 720p. For multi-modal inputs as reference, its current open platform supports up to 3 video clips, 9 images, and 3 audio clips. In addition, we provide Seedance 2.0 Fast version, an accelerated variant of Seedance 2.0 designed to boost generation speed for low-latency scenarios. Seedance 2.0 has delivered significant improvements to its foundational generation capabilities and multi-modal generation performance, bringing an enhanced creative experience for end users.

核心贡献 · Key contributions

局限 · Limitations

论文章节 · Sections(共 6)

阅读逐段中英对照全文 →