5 papers
UltraVR: A Diagnostic Ultra-Resolution Image-VQA Benchmark for Evidence-Grounded Reasoning
Gexin Huang, Yanting Yang, Myeongkyun Kang +6
Vision-language models (VLMs) excel on visual question answering and multimodal reasoning benchmarks. Yet their capability on ultra-resolution images - where critical evidence is t…
Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models
Gexin Huang, Anqi Li, Yusheng Tan +4
Pathology foundation models (FMs) have become central to computational histopathology, offering strong transfer performance across a wide range of diagnostic and prognostic tasks.…
See-in-Pairs: Reference Image-Guided Comparative Vision-Language Models for Medical Diagnosis
Ruinan Jin, Gexin Huang, Xinwei Shen +3
Medical image diagnosis is challenging because many diseases resemble normal anatomy and exhibit substantial interpatient variability. Clinicians routinely rely on comparative diag…
Identity Preserving Latent Diffusion for Brain Aging Modeling
Gexin Huang, Zhangsihao Yang, Yalin Wang +3
Structural and appearance changes in brain imaging over time are crucial indicators of neurodevelopment and neurodegeneration. The rapid advancement of large-scale generative model…
Interactive Tumor Progression Modeling via Sketch-Based Image Editing
Gexin Huang, Ruinan Jin, Yucheng Tang +4
Accurately visualizing and editing tumor progression in medical imaging is crucial for diagnosis, treatment planning, and clinical communication. To address the challenges of subje…