3 papers
cs.CV2025
Layer-Aware Video Composition via Split-then-Merge
Ozgur Kara, Yujia Chen, Ming-Hsuan Yang +3
We present Split-then-Merge (StM), a novel framework designed to enhance control in generative video composition and address its data scarcity problem. Unlike conventional methods…
cs.CV2025
SEAL: Semantic Attention Learning for Long Video Representation
Lan Wang, Yujia Chen, Du Tran +2
Long video understanding presents challenges due to the inherent high computational complexity and redundant temporal information. An effective representation for long videos must…
cs.LG2024
Semantic Image Inversion and Editing using Rectified Stochastic Differential Equations
Litu Rout, Yujia Chen, Nataniel Ruiz +3
Generative models transform random noise into images; their inversion aims to transform images back to structured noise for recovery and editing. This paper addresses two key tasks…