Signals 4 · free daily AI digest

SenseNova-U1.5: Towards Native Unified Visual Intelligence

Paper recorded by Signals 4 on 2026-09-10 in cs.CV. Abstract reproduced from arXiv; link to the original below.

Published 2026-09-10 on arXiv · recorded by Signals 4 on 2026-09-11

Category: cs.CV · 计算机视觉 · first seen 2026-09-11

Abstract

We launch SenseNova-U1.5, an 8B-MoT native unified multimodal model that understands, reasons about, and generates visual content within an encoder-free and VAE-free architecture. We strengthen its visual interface through spatially coherent patch reconstruction and scale its training with carefully curated generation and editing data, improved task formulation, structural prompt enhancement, and

Read on arXiv →

#79 most recent of 237 cs.CV papers we have recorded · ↑ newer: MGAvatar: Mesh-Bound Gaussians for Head Avatar Geometry and Appearance · ↓ older: Caption-once, Frames-on-Demand: Visual-Need Routing for Budget-Aware A
Cite this page: SenseNova-U1.5: Towards Native Unified Visual Intelligence: the #79 most recent of 237 cs.CV papers we have recorded (as of 2026-09-10). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/sensenova-u1-5-towards-native-unified-visual-intelligence.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable: papers.json
Related: More cs.CV papers · arXiv signals · All papers · Today in AI
Get 4 AI signals a day by email — free.
Subscribe free → See all plans →
Get 4 AI signals a day by email — free
All models · All repos · By company · Daily editions