5 papers
StoryMem: Multi-shot Long Video Storytelling with Memory
Kaiwen Zhang, Liming Jiang, Angtian Wang +6
Visual storytelling requires generating multi-shot videos with cinematic quality and long-range consistency. Inspired by human memory, we propose StoryMem, a paradigm that reformul…
Learning Joint ID-Textual Representation for ID-Preserving Image Synthesis
Zichuan Liu, Liming Jiang, Qing Yan +3
We propose a novel framework for ID-preserving generation using a multi-modal encoding strategy rather than injecting identity features via adapters into pre-trained models. Our me…
Flux Already Knows -- Activating Subject-Driven Image Generation without Training
Hao Kang, Stathi Fotiadis, Liming Jiang +5
We propose a simple yet effective zero-shot framework for subject-driven image generation using a vanilla Flux model. By framing the task as grid-based image completion and simply…
NTIRE 2025 Challenge on Event-Based Image Deblurring: Methods and Results
Lei Sun, Andrea Alfarano, Peiqi Duan +85
This paper presents an overview of NTIRE 2025 the First Challenge on Event-Based Image Deblurring, detailing the proposed methodologies and corresponding results. The primary goal…
InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
Liming Jiang, Qing Yan, Yumin Jia +3
Achieving flexible and high-fidelity identity-preserved image generation remains formidable, particularly with advanced Diffusion Transformers (DiTs) like FLUX. We introduce Infini…