3 papers
cs.CV2026
MedGround: Bridging the Evidence Gap in Medical Vision-Language Models with Verified Grounding Data
Mengmeng Zhang, Xiaoping Wu, Hao Luo +2
Vision-Language Models (VLMs) can generate convincing clinical narratives, yet frequently struggle to visually ground their statements. We posit this limitation arises from the sca…
cs.LG2025
Context-Aware Probabilistic Modeling with LLM for Multimodal Time Series Forecasting
Yueyang Yao, Jiajun Li, Xingyuan Dai +4
Time series forecasting is important for applications spanning energy markets, climate analysis, and traffic management. However, existing methods struggle to effectively integrate…
cs.CV2025
Hierarchical Self-Prompting SAM: A Prompt-Free Medical Image Segmentation Framework
Mengmeng Zhang, Xingyuan Dai, Yicheng Sun +6
Although the Segment Anything Model (SAM) is highly effective in natural image segmentation, it requires dependencies on prompts, which limits its applicability to medical imaging…