2 papers
cs.LG2026
Compression as Adaptation: Implicit Visual Representation with Diffusion Foundation Models
Zongyu Guo, Jiajun He, Zhaoyang Jia +6
Modern visual generative models acquire rich visual knowledge through large-scale training, yet existing visual representations (such as pixels, latents, or tokens) remain external…
cs.CV2026
Narrative Weaver: Towards Controllable Long-Range Visual Consistency with Multi-Modal Conditioning
Zhengjian Yao, Yongzhi Li, Xinyuan Gao +3
We present "Narrative Weaver", a novel framework that addresses a fundamental challenge in generative AI: achieving multi-modal controllable, long-range, and consistent visual cont…