🛰️ Daily AI Frontier
‹ back to 2026-09-22

阿里公布全模态模型新进展,Qwen4和下代视频模型均在训练中

Industry & News Multimodal & Generative

Ranking

Overall 68
Content 75
Popularity N/A

No observed public metrics; popularity remains neutral/archived.

Representative image for 阿里公布全模态模型新进展,Qwen4和下代视频模型均在训练中

Merged summary

TL;DR - Alibaba announced broad updates across its Qwen and Wan model families, including Qwen4 training, a next-generation video model due in November, and new image, audio, translation, and world models. The roadmap emphasizes unified multimodal intelligence, lower compute costs, larger models, and partially autonomous model improvement.

  • Qwen4 is training on a new architecture, while Qwen4.5 and Qwen5 are planned to scale toward 5–10 trillion total parameters.
  • Qwen3.8-Max reportedly completed 33 effective self-improvement iterations over a month without human participation, autonomously constructing data, experiments, and training workflows.
  • Qwen3.8-Flash cuts training costs by nearly 90%, while autonomous adaptation of its SGLang stack to a new GPU reportedly increased single-instance inference throughput by 96%.
  • Alibaba also unveiled Qwen3.8-Omni-Flash, Qwen-Audio-3.1, Qwen-Image-3.1, and HappyOyster 2.0 Preview; its next video model targets longer, more controllable, narratively coherent generation.

Sources (1)

阿里公布全模态模型新进展,Qwen4和下代视频模型均在训练中

量子位 量子位的朋友们 2026-09-22
Public signals N/A
Providers: Hugging Face · N/A OpenAlex · N/A Publisher · N/A Semantic Scholar · N/A X · N/A Fetched 2026-09-26 14:14:38.758373 UTC

TL;DR - Alibaba announced broad updates across its Qwen and Wan model families, including Qwen4 training, a next-generation video model due in November, and new image, audio, translation, and world models. The roadmap emphasizes unified multimodal intelligence, lower compute costs, larger models, and partially autonomous model improvement.

  • Qwen4 is training on a new architecture, while Qwen4.5 and Qwen5 are planned to scale toward 5–10 trillion total parameters.
  • Qwen3.8-Max reportedly completed 33 effective self-improvement iterations over a month without human participation, autonomously constructing data, experiments, and training workflows.
  • Qwen3.8-Flash cuts training costs by nearly 90%, while autonomous adaptation of its SGLang stack to a new GPU reportedly increased single-instance inference throughput by 96%.
  • Alibaba also unveiled Qwen3.8-Omni-Flash, Qwen-Audio-3.1, Qwen-Image-3.1, and HappyOyster 2.0 Preview; its next video model targets longer, more controllable, narratively coherent generation.
item →