SenseNova U1.5 Lite正式版发布:支持超长指令,解锁原生4K真实视觉创作流
TL;DR - SenseTime has open-sourced SenseNova U1.5 Lite, an 8B unified multimodal model designed for production-oriented image generation and editing. It targets complex prompt adherence, native 4K output, precise text and layout rendering, and controllable edits while remaining deployable on a single GPU.
- Supports 3–4K-character prompts with constraints spanning subjects, spatial relationships, text, layout, and style.
- Adds native 4K generation, multi-reference composition, bounding-box and visual-marker controls, and preservation of non-edited regions.
- Uses the NEO-unify architecture and MOPD distillation to consolidate specialized training experts into one 8B model without an inference-time router.
- Strengthens task-oriented reinforcement-learning post-training around instruction compliance, visual quality, and editing preservation.