开源国产8B模型,比肩闭源Image 2了!
Ranking
Overall
68
Content
75
Popularity
N/A
No observed public metrics; popularity remains neutral/archived.
Merged summary
TL;DR - SenseTime has released the open-source SenseNova U1.5 Lite, an 8B unified model for image generation and editing with native 4K output. The release targets production-ready visual workflows through stronger instruction following, layout and text rendering, precise editing, and preservation of unmodified content.
- The model supports 3,000–4,000-character prompts, multiple image references, region annotations, complex layouts, and Chinese and English text.
- Its NEO-unify architecture combines visual understanding, image generation, and editing in one model without an external expert router.
- Specialized experts for text rendering, aesthetics, and editing are consolidated into the 8B model through multi-teacher online policy distillation (MOPD).
- Post-training emphasizes instruction adherence, visual quality, and edit preservation; the article claims performance comparable to closed-source GPT-Image-2 in selected layout and editing tasks.
Sources (1)
开源国产8B模型,比肩闭源Image 2了!
Public signals
N/A
TL;DR - SenseTime has released the open-source SenseNova U1.5 Lite, an 8B unified model for image generation and editing with native 4K output. The release targets production-ready visual workflows through stronger instruction following, layout and text rendering, precise editing, and preservation of unmodified content.
- The model supports 3,000–4,000-character prompts, multiple image references, region annotations, complex layouts, and Chinese and English text.
- Its NEO-unify architecture combines visual understanding, image generation, and editing in one model without an external expert router.
- Specialized experts for text rendering, aesthetics, and editing are consolidated into the 8B model through multi-teacher online policy distillation (MOPD).
- Post-training emphasizes instruction adherence, visual quality, and edit preservation; the article claims performance comparable to closed-source GPT-Image-2 in selected layout and editing tasks.