3 papers
cs.CV2026
BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing
Peilin Xiong, Honghui Yuan, Junwen Chen +1
Coarse-mask local image editing asks a model to modify a user-indicated region while preserving the surrounding scene. In practice, however, rough masks often become unintended sha…
cs.CV2026
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
Junwen Chen, Peilin Xiong, Keiji Yanai
Recent human-object interaction detection (HOID) methods highly require prior knowledge from vision-language models (VLMs) to enhance the interaction recognition capabilities. The…
cs.CV2025
PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing
Peilin Xiong, Junwen Chen, Honghui Yuan +1
Localized subject-driven image editing aims to seamlessly integrate user-specified objects into target scenes. As generative models continue to scale, training becomes increasingly…