3 papers
cs.CV2026
ASTRA: Let Arbitrary Subjects Transform in Video Editing
Fei Shen, Weihao Xu, Rui Yan +4
While existing video editing methods excel with single subjects, they struggle in dense, multi-subject scenes, frequently suffering from attention dilution and mask boundary entang…
cs.CV2025
R-Genie: Reasoning-Guided Generative Image Editing
Dong Zhang, Lingfeng He, Rui Yan +2
While recent advances in image editing have enabled impressive visual synthesis capabilities, current methods remain constrained by explicit textual instructions and limited editin…
cs.CV2025
Memory Efficient Transformer Adapter for Dense Predictions
Dong Zhang, Rui Yan, Pingcheng Dong +1
While current Vision Transformer (ViT) adapter methods have shown promising accuracy, their inference speed is implicitly hindered by inefficient memory access operations, e.g., st…