activity
20242026
collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2026

CAER: Conflict-Aware Evidence Routing with Dual Prefix Experts for Multimodal Large Language Models

Zixuan Liu, Juntao Cai, Xiaoxu Cai +2

Multimodal Large Language Models (MLLMs) have demonstrated remarkable capabilities in multimodal understanding and generation. However, when textual inputs conflict with visual evi…

cs.CV2026

Morphology-Aware Multimodal Representation Learning for Insect Phylogenetic Reconstruction

Zixuan Liu, Kaijie Yu, Chun He +5

Morphological traits provide important evidence for phylogenetic reconstruction and evolutionary relationship analysis. Recent image-based approaches have introduced deep learning,…

cs.CV2026

Directed Ordinal Diffusion Regularization for Progression-Aware Diabetic Retinopathy Grading

Huangwei Chen, Junhao Jia, Ruocheng Li +7

Diabetic Retinopathy (DR) progresses as a continuous and irreversible deterioration of the retina, following a well-defined clinical trajectory from mild to severe stages. However,…

cs.CV2026

VasGuideNet: Vascular Topology-Guided Couinaud Liver Segmentation with Structural Contrastive Loss

Chaojie Shen, Jingjun Gu, Zihao Zhao +4

Accurate Couinaud liver segmentation is critical for preoperative surgical planning and tumor localization.However, existing methods primarily rely on image intensity and spatial l…

cs.CV2025

One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning

Chunpeng Zhou, Qianqian Shen, Zhi Yu +2

Recent advancements in fine-tuning Vision-Language Foundation Models (VLMs) have garnered significant attention for their effectiveness in downstream few-shot learning tasks.While…

cs.CV2025

MetaNeRV: Meta Neural Representations for Videos with Spatial-Temporal Guidance

Jialong Guo, Ke liu, Jiangchao Yao +3

Neural Representations for Videos (NeRV) has emerged as a promising implicit neural representation (INR) approach for video analysis, which represents videos as neural networks wit…