Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
Chengxu Yang, Jingling Yuan, Chuang Hu +1
Multimodal Large Language Models often suffer from object hallucination. While existing research utilizes attention enhancement and visual retracing, we find these works lack suffi…
cs.CV2025
RAPNet: A Receptive-Field Adaptive Convolutional Neural Network for Pansharpening
Tao Tang, Chengxu Yang
Pansharpening refers to the process of integrating a high resolution panchromatic (PAN) image with a lower resolution multispectral (MS) image to generate a fused product, which is…