activity
20242026
collaborators

16 papers

cs.CL2026

Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate

Zhixiang Lu, Jionglong Su

Multimodal Large Language Models (MLLMs) in healthcare suffer from severe confirmation bias, often hallucinating visual details to support initial, potentially erroneous diagnostic…

cs.CV2026

Attention-Guided Flow-Matching for Sparse 3D Geological Generation

Zhixiang Lu, Mengqi Han, Peixin Guo +4

Constructing high-resolution 3D geological models from sparse 1D borehole and 2D surface data is a highly ill-posed inverse problem. Traditional heuristic and implicit modeling met…

cs.CV2026

Semantic-Topological Graph Reasoning for Language-Guided Pulmonary Screening

Chenyu Xue, Yiran Liu, Mian Zhou +2

Medical image segmentation driven by free-text clinical instructions is a critical frontier in computer-aided diagnosis. However, existing multimodal and foundation models struggle…

cs.CV2026

ConFoThinking: Consolidated Focused Attention Driven Thinking for Visual Question Answering

Zhaodong Wu, Haochen Xue, Qi Cao +5

Thinking with Images improves fine-grained VQA for MLLMs by emphasizing visual cues. However, tool-augmented methods depend on the capacity of grounding, which remains unreliable f…

cs.CL2025

PRISM: A Personality-Driven Multi-Agent Framework for Social Media Simulation

Zhixiang Lu, Xueyuan Deng, Yiran Liu +4

Traditional agent-based models (ABMs) of opinion dynamics often fail to capture the psychological heterogeneity driving online polarization due to simplistic homogeneity assumption…

cs.CV2025

SAM-DCE: Addressing Token Uniformity and Semantic Over-Smoothing in Medical Segmentation

Yingzhen Hu, Yiheng Zhong, Ruobing Li +5

The Segment Anything Model (SAM) demonstrates impressive zero-shot segmentation ability on natural images but encounters difficulties in medical imaging due to domain shifts, anato…