🛰️ Daily AI Frontier
‹ back to 2026-08-24

阿里视频大模型Wan3.0正式上线,行业评价“稳定、真实、有质感”

Industry & News Multimodal & Generative 🔗 2 sources

Ranking

Overall 64
Content 70
Popularity N/A

No observed public metrics; popularity remains neutral/archived.

Representative image for 阿里视频大模型Wan3.0正式上线,行业评价“稳定、真实、有质感”

Merged summary

TL;DR — 阿里巴巴正式推出视频生成模型 Wan3.0,主打稳定性、真实性与画面质感,可生成最长 30 秒的视频,并支持文档作为输入。其面向影视、广告、旅游、音乐视频及创作者平台等生产场景,旨在提升 AI 视频的规模化实用性。

  • 在较长视频中保持人物、物体、音频、空间关系和视觉风格的一致性。
  • 支持 DOC、XLS、PPT、PDF 和 Markdown 等文档输入,也可结合参考媒体进行视频创作。
  • 强化人脸、皮肤、材质、光影、动作和微表情等细节,追求自然表达、连贯运动及符合物理规律的呈现。
  • 已通过阿里云百炼、通义千问相关平台、万相网站及移动端等渠道上线,并提供商业 API 和部分第三方服务接入。

注: 量子位更强调可规模化落地及创作者平台应用,雷峰网更侧重生产工作流与具体视觉细节;两者对核心能力和目标场景的描述基本一致。

Sources (2)

阿里视频大模型Wan3.0正式上线,行业评价“稳定、真实、有质感”

量子位 量子位的朋友们 2026-08-24
Public signals N/A
Providers: Hugging Face · N/A OpenAlex · N/A Publisher · N/A Semantic Scholar · N/A X · N/A Fetched 2026-09-23 14:19:19.989290 UTC

TL;DR - Alibaba launched Wan3.0, a video-generation model focused on production-ready consistency, realism, and visual quality. It generates videos up to 30 seconds and accepts documents as inputs, targeting scalable use in film, advertising, creator platforms, tourism, and music videos.

  • Maintains characters, objects, audio, spatial relationships, and visual style across longer sequences.
  • Supports DOC, XLS, PPT, PDF, and Markdown inputs for reference-driven video creation.
  • Emphasizes realistic human details, natural expressions, coherent motion, and consistent physical and lighting properties.
  • Available through Alibaba Cloud Model Studio, Qwen platforms, Wanxiang, and partner services, with API access offered commercially.
item →

阿里视频大模型Wan3.0正式上线,行业评价“稳定、真实、有质感”

雷峰网 (AI科技评论) 2026-08-24
Public signals N/A
Providers: Hugging Face · N/A OpenAlex · N/A Publisher · N/A Semantic Scholar · N/A X · N/A Fetched 2026-09-23 14:19:18.810942 UTC

TL;DR - Alibaba launched Wan3.0, a video-generation model aimed at production workflows, with up to 30-second generation, document-based inputs, and stronger temporal and visual consistency. Its emphasis on stable, realistic output could make AI video more practical for film, advertising, tourism, and music-video production.

  • Generates videos up to 30 seconds while maintaining consistency across characters, objects, audio, spatial relationships, and visual style.
  • Adds support for DOC, XLS, PPT, PDF, and Markdown inputs alongside reference media.
  • Targets realistic rendering of faces, skin, motion, lighting, materials, micro-expressions, and other physical details.
  • Available through Alibaba Cloud Model Studio, Qwen platforms, Wan’s website, mobile apps, APIs, and several third-party services.
item →