1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026
PaceVGGT: Pre-Alternating-Attention Token Pruning for Visual Geometry Transformers
Haotang Li, Zhenyu Qi, Shaohan Henry Wang +5
Visual Geometry Transformer (VGGT) is a strong feed-forward model for multiple 3D tasks, but its Alternating-Attention (AA) stack scales quadratically in the total token count, mak…
cs.CV2024★ 1 cited
CLII: Visual-Text Inpainting via Cross-Modal Predictive Interaction
Liang Zhao, Qing Guo, Xiaoguang Li +1
Image inpainting aims to fill missing pixels in damaged images and has achieved significant progress with cut-edging learning techniques. Nevertheless, state-of-the-art inpainting…
cs.CV2023
SAIR: Learning Semantic-aware Implicit Representation
Canyu Zhang, Xiaoguang Li, Qing Guo +1
Implicit representation of an image can map arbitrary coordinates in the continuous domain to their corresponding color values, presenting a powerful capability for image reconstru…