Skip to content
AI Atlas
PaperActive

SenseNova-U1.5: Towards Native Unified Visual Intelligence

arxiv.org/abs/2609.11929

Updated 41 min ago · first seen 11 Sept 2026

paper_01M294H2N5RNFQDAHHFE1QJJEA

Published
11 Sept 2026
T1 · 50 min ago
arXiv
2609.11929
T1 · 50 min ago
Category
cs.CV
T1 · 50 min ago

Abstract

We launch SenseNova-U1.5, an 8B-MoT native unified multimodal model that understands, reasons about, and generates visual content within an encoder-free and VAE-free architecture. We strengthen its visual interface through spatially coherent patch reconstruction and scale its training with carefully curated generation and editing data, improved task formulation, structural prompt enhancement, and native resolutions of up to 4K. For post-training, we optimize specialized experts for visual aesthetics, bilingual text rendering, infographic generation, and image editing, and consolidate their capabilities through multi-expert on-policy distillation. Across extensive evaluations, SenseNova-U1.5 largely advances image fidelity, text rendering, complex composition, multi-reference editing, and interleaved generation, while improving instruction following and preserving subject identity, geometry, and unmodified regions. Despite limited exposure to structured formats in its generation data, SenseNova-U1.5 generalizes effectively to long, complex, and structured visual instructions, further proving that multimodal understanding can transfer to visual planning and creation. Together, these findings position native unified modelling as a promising path towards systems that perceive, reason and create within a fully end-to-end framework. We will open-source training code, including supervised fine-tuning, reinforcement learning, and on-policy distillation.

Authors 65

Haiwen Diao, Jiahao Wang, Chenjing Ding, Hanming Deng, Jiangnan Chen, Ruixi Zhang, Ruohui Wang, Wenwen Tong, Xiangyu Fan, Yubo Wang, Yue Zhu, Yuwei Niu, Zhengqi Bai, Zhiqian Lin, Zhitao Yang, Zhongang Cai, Bo Yang, Chen Feng, Chengguang Lv, Guangjia Liu, Guanlin Wang, Hanyu Zhang, Haojia Yu, Hongcan Xiao, Hongli Wang, Huan Wu, Huaping Zhong, Jian Fang, Jianan Fan, Jiaqi Li, Jiefan Lu, Jing Zuo, Jingcheng Ni, Junxiang Xu, Linjun Dai, Mutian Xu, Peishen Yan, Penghao Wu, Ruijie Mao, Ruisi Wang, Shihao Bai, Shuang Yang, Shuya Yang, Shuyan Zheng, Silei Wu, Siying Li, Tao Chu, Tianbo Zhong, Tongxi Zhou, Weichao Luo, Weichen Fan, Wenhao Jia, Wenjie Gao, Xiangli Kong, Yan Li, Yang Yong, Zimo Wen, Zixuan Qian, Wenxiu Sun, Ruihao Gong, Quan Wang, Lewei Lu, Lei Yang, Ziwei Liu, Dahua Lin

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

arXiv id
2609.11929

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Categories
cs.CV

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Hf paper url
https://huggingface.co/papers/2609.11929

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 44 min agomedium

Hf comments
1

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 41 min agomedium

Upvotes
128

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 41 min agomedium

PDF

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Primary category
cs.CV

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

12

Source tiers

T1T29 / 3

Freshest observation

41 min ago

Conflicts

4 flagged