collaborators

6 papers

cs.CV2026

DeepInv: A Novel Self-supervised Learning Approach for Fast and Accurate Diffusion Inversion

Ziyue Zhang, Luxi Lin, Xiaolin Hu +4

Diffusion inversion is a task of recovering the noise of an image in a diffusion model, which is vital for controllable diffusion image editing. At present, diffusion inversion sti…

cs.AR2025

Focus: A Streaming Concentration Architecture for Efficient Vision-Language Models

Chiyue Wei, Cong Guo, Junyao Zhang +8

Vision-Language Models (VLMs) have demonstrated strong performance on tasks such as video captioning and visual question answering. However, their growing scale and video-level inp…

cs.CV2025

ObjectAdd: Adding Objects into Image via a Training-Free Diffusion Modification Fashion

Ziyue Zhang, Mingbao Lin, Quanjian Song +2

We introduce ObjectAdd, a training-free diffusion modification method to add user-expected objects into user-specified area. The motive of ObjectAdd stems from: first, describing e…

eess.SP2025

EMind: A Foundation Model for Multi-task Electromagnetic Signals Understanding

Luqing Luo, Wenjin Gui, Yunfei Liu +11

Deep understanding of electromagnetic signals is fundamental to dynamic spectrum management, intelligent transportation, autonomous driving and unmanned vehicle perception. The fie…

cs.CV2025

EasyInv: Toward Fast and Better DDIM Inversion

Ziyue Zhang, Mingbao Lin, Shuicheng Yan +1

This paper introduces EasyInv, an easy yet novel approach that significantly advances the field of DDIM Inversion by addressing the inherent inefficiencies and performance limitati…

cs.CV2025

LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation

Quanjian Song, Zhihang Lin, Zhanpeng Zeng +3

Existing camera motion-controlled video generation methods face computational bottlenecks in fine-tuning and inference. This paper proposes LightMotion, a light and tuning-free met…