collaborators

9 papers

cs.CV2025

LiMT: A Multi-task Liver Image Benchmark Dataset

Zhe Liu, Kai Han, Siqi Ma +10

Computer-aided diagnosis (CAD) technology can assist clinicians in evaluating liver lesions and intervening with treatment in time. Although CAD technology has advanced in recent y…

cs.CV2025

Frequency Domain Unlocks New Perspectives for Abdominal Medical Image Segmentation

Kai Han, Siqi Ma, Chengxuan Qian +4

Accurate segmentation of tumors and adjacent normal tissues in medical images is essential for surgical planning and tumor staging. Although foundation models generally perform wel…

cs.CV2025

Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools

Zhenlong Yuan, Xiangyan Qu, Chengxuan Qian +8

Multimodal large language models (MLLMs) have demonstrated remarkable potential in bridging visual and textual reasoning, yet their reliance on text-centric priors often limits the…

cs.CV2025

CLIMD: A Curriculum Learning Framework for Imbalanced Multimodal Diagnosis

Kai Han, Chongwen Lyu, Lele Ma +5

Clinicians usually combine information from multiple sources to achieve the most accurate diagnosis, and this has sparked increasing interest in leveraging multimodal deep learning…

cs.CV2025

HAIF-GS: Hierarchical and Induced Flow-Guided Gaussian Splatting for Dynamic Scene

Jianing Chen, Zehao Li, Yujun Cai +7

Reconstructing dynamic 3D scenes from monocular videos remains a fundamental challenge in 3D vision. While 3D Gaussian Splatting (3DGS) achieves real-time rendering in static setti…

cs.CV2025

Adaptive Label Correction for Robust Medical Image Segmentation with Noisy Labels

Chengxuan Qian, Kai Han, Jianxia Ding +4

Deep learning has shown remarkable success in medical image analysis, but its reliance on large volumes of high-quality labeled data limits its applicability. While noisy labeled d…