🛰️ Daily AI Frontier
‹ back to 2026-09-22

阿里公布全模态模型新进展,Qwen4和下代视频模型均在训练中

量子位 Multimodal & Generative 量子位的朋友们 2026-09-22
Representative image for 阿里公布全模态模型新进展,Qwen4和下代视频模型均在训练中

TL;DR - Alibaba announced broad updates across its Qwen and Wan model families, including Qwen4 training, a next-generation video model due in November, and new image, audio, translation, and world models. The roadmap emphasizes unified multimodal intelligence, lower compute costs, larger models, and partially autonomous model improvement.

  • Qwen4 is training on a new architecture, while Qwen4.5 and Qwen5 are planned to scale toward 5–10 trillion total parameters.
  • Qwen3.8-Max reportedly completed 33 effective self-improvement iterations over a month without human participation, autonomously constructing data, experiments, and training workflows.
  • Qwen3.8-Flash cuts training costs by nearly 90%, while autonomous adaptation of its SGLang stack to a new GPU reportedly increased single-instance inference throughput by 96%.
  • Alibaba also unveiled Qwen3.8-Omni-Flash, Qwen-Audio-3.1, Qwen-Image-3.1, and HappyOyster 2.0 Preview; its next video model targets longer, more controllable, narratively coherent generation.

view merged work →