works on

From the 1 of 10 linked papers with an AI index.

activity
20242026
collaborators

10 papers

cs.CV2026

Wonder: Video World Model Done Better

Jiacong Xu, Hanwen Jiang, Zhixin Shu +3

Wonder is a video world model that lets users explore a generated scene in real time by moving a virtual camera, using a dense coordinate conditioning and a sparse attention memory…

cs.CV2026

TokenLight: Precise Lighting Control in Images using Attribute Tokens

Sumit Chaturvedi, Yannick Hold-Geoffroy, Mengwei Ren +5

This paper presents a method for image relighting that enables precise and continuous control over multiple illumination attributes in a photograph. We formulate relighting as a co…

cs.CV2025

Endless World: Real-Time 3D-Aware Long Video Generation

Ke Zhang, Yiqun Mei, Jiacong Xu +1

Producing long, coherent video sequences with stable 3D structure remains a major challenge, particularly in streaming scenarios. Motivated by this, we introduce Endless World, a r…

cs.CV2025

RELIC: Interactive Video World Model with Long-Horizon Memory

Yicong Hong, Yiqun Mei, Chongjian Ge +11

A truly interactive world model requires three key ingredients: real-time long-horizon streaming, consistent spatial memory, and precise user control. However, most existing approa…

cs.CV2025

Think Before You Diffuse: Infusing Physical Rules into Video Diffusion

Ke Zhang, Cihan Xiao, Jiacong Xu +2

Recent video diffusion models have demonstrated their great capability in generating visually-pleasing results, while synthesizing the correct physical effects in generated videos…

cs.CV2025

FreeViS: Training-free Video Stylization with Inconsistent References

Jiacong Xu, Yiqun Mei, Ke Zhang +1

Video stylization plays a key role in content creation, but it remains a challenging problem. Naïvely applying image stylization frame-by-frame hurts temporal consistency and redu…