17 citations · 31 across the 18 of their papers we have counts for
26 papers · 1 filter
LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression
Bowen Yuan, Zijian Wang, Yadan Luo +2
Large vision-language models (LVLMs) exhibit strong reasoning ability but suffer from visual forgetting during long-horizon decoding, where attention progressively drifts away from…
Divide-and-Conquer Approach to Holistic Cognition in High-Similarity Contexts with Limited Data
Shijie Wang, Zijian Wang, Yadan Luo +3
Ultra-fine-grained visual categorization (Ultra-FGVC) aims to classify highly similar subcategories within fine-grained objects using limited training samples. However, holistic ye…
Geometry-Guided Self-Supervision for Ultra-Fine-Grained Recognition with Limited Data
Shijie Wang, Yadan Luo, Zijian Wang +3
This paper investigates the intrinsic geometrical features of highly similar objects and introduces a general self-supervised framework called the Geometric Attribute Exploration N…
Language-driven Fine-grained Retrieval
Shijie Wang, Xin Yu, Yadan Luo +3
Existing fine-grained image retrieval (FGIR) methods learn discriminative embeddings by adopting semantically sparse one-hot labels derived from category names as supervision. Whil…
Distributed Zero-Shot Learning for Visual Recognition
Zhi Chen, Yadan Luo, Zi Huang +3
In this paper, we propose a Distributed Zero-Shot Learning (DistZSL) framework that can fully exploit decentralized data to learn an effective model for unseen classes. Considering…
Latent Refinement via Flow Matching for Training-free Linear Inverse Problem Solving
Hossein Askari, Yadan Luo, Hongfu Sun +1
Recent advances in inverse problem solving have increasingly adopted flow priors over diffusion models due to their ability to construct straight probability paths from noise to da…