3 papers
cs.CV2026
Counterfactual Anatomy-guided Spatial-Temporal Decoding for Annotation-Free Hallucination Mitigation in Medical VLMs
Yifan Lu, Adinath Dukre, Abhijit Das +4
Medical vision-language models (Med-VLMs) have demonstrated strong performance on medical visual question answering, yet they remain prone to hallucination, generating clinically u…
cs.LG2026
CHIPS: Efficient CLIP Adaptation via Curvature-aware Hybrid Influence-based Data Selection
Xinlin Zhuang, Yichen Li, Xiwei Liu +11
Adapting CLIP to vertical domains is typically approached by novel fine-tuning strategies or by continual pre-training (CPT) on large domain-specific datasets. Yet, data itself rem…
cs.LG2025
HeSRN: Representation Learning On Heterogeneous Graphs via Slot-Aware Retentive Network
Yifan Lu, Ziyun Zou, Belal Alsinglawi +6
Graph Transformers have recently achieved remarkable progress in graph representation learning by capturing long-range dependencies through self-attention. However, their quadratic…