5 citations · 12 across the 10 of their papers we have counts for
4 papers · 1 filter
SeVeR: Selective Visual Exposure and Retrieval for 3D Medical Image Question Answering
Yaojun Hu, Danyang Tu, Yang Liu +8
Volumetric medical VQA requires reasoning over long and redundant 3D visual token sequences, especially in multi-sequence MRI where complementary modalities provide diverse diagnos…
D3O: Dynamic Distribution Distillation for Ordinal Regression
Chunlai Dong, Yaojun Hu, Yuyang Xu +2
Ordinal regression is widely used in scenarios where labels are discrete yet inherently ordered. In practice, however, ordinal labels are often obtained by discretizing underlying…
HyperVLP: Enhancing Hierarchical Surgical Video-Language Pre-training in Hyperbolic Space
Yaojun Hu, Kun Yuan, Nassir Navab +3
Surgical vision-language foundation models typically adopt educational materials, such as surgical lecture videos, to transfer surgical knowledge encoded in language into visual re…
BreastGPT: A Multimodal Large Language Model for the Full Spectrum of Breast Cancer Clinical Routine
Yang Liu, Jiajin Zhang, Danyang Tu +8
Breast cancer remains a leading cause of cancer-related mortality among women. Its clinical management requires multimodal reasoning across a clinical workflow that spans \textit{s…