SenseNova-U1.5 Technical Report: VAE-free Native 4K Generation with Spatial Patch Reconstruction
This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.
SenseNova has released the technical report for SenseNova-U1.5, an 8B native unified model for image understanding, generation, and editing. The model does not rely on an external visual encoder or VAE. Images are mapped directly into visual tokens, with each token representing a 32×32-pixel region…
Read the full story at r/StableDiffusion ↗
Timeline · 1 report
- 2026-09-16 13:09 · r/StableDiffusion
SenseNova-U1.5 Technical Report: VAE-free Native 4K Generation with Spatial Patch Reconstruction