3 papers
cs.CV2026
Illustrator's Depth: Monocular Layer Index Prediction for Image Decomposition
Nissim Maruani, Peiying Zhang, Siddhartha Chaudhuri +6
We introduce Illustrator's Depth, a novel definition of depth that addresses a key challenge in digital content creation: decomposing flat images into editable, ordered layers. Ins…
cs.CV2025
DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance
Peiying Zhang, Nanxuan Zhao, Matthew Fisher +3
Recent vision-language model (VLM)-based approaches have achieved impressive results on SVG generation. However, because they generate only text and lack visual signals during deco…
cs.CV2025
RELIC: Interactive Video World Model with Long-Horizon Memory
Yicong Hong, Yiqun Mei, Chongjian Ge +11
A truly interactive world model requires three key ingredients: real-time long-horizon streaming, consistent spatial memory, and precise user control. However, most existing approa…