Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
A World Model of Radiologist Reading for Medical Image Representation Learning
Yiwei Li, Zihao Wu, Huaqin Zhao +5
Radiologist eye-tracking data provide a rich record of how experts search, compare, and accumulate evidence during image reading; yet, existing methods exploit this signal only par…
cs.CV2025
AdCare-VLM: Towards a Unified and Pre-aligned Latent Representation for Healthcare Video Understanding
Md Asaduzzaman Jabin, Hanqi Jiang, Yiwei Li +4
Chronic diseases, including diabetes, hypertension, asthma, HIV-AIDS, epilepsy, and tuberculosis, necessitate rigorous adherence to medication to avert disease progression, manage…
cs.CV2025
Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models
Yifan Xu, Chao Zhang, Hanqi Jiang +6
Advancements in foundation models have made it possible to conduct applications in various downstream tasks. Especially, the new era has witnessed a remarkable capability to extend…