1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026★ 1 cited
SkinCLIP-VL: Consistency-Aware Vision-Language Learning for Multimodal Skin Cancer Diagnosis
Zhixiang Lu, Shijie Xu, Kaicheng Yan +6
The deployment of vision-language models (VLMs) in dermatology is hindered by the trilemma of high computational costs, extreme data scarcity, and the black-box nature of deep lear…
cs.CV2026
Causal-SAM-LLM: Large Language Models as Causal Reasoners for Robust Medical Segmentation
Tao Tang, Shijie Xu, Jionglong Su +1
The clinical utility of deep learning models for medical image segmentation is severely constrained by their inability to generalize to unseen domains. This failure is often rooted…
cs.CV2025
Xiaomi MiMo-VL-Miloco Technical Report
Jiaze Li, Jingyang Chen, Yuxun Qu +9
We open-source MiMo-VL-Miloco-7B and its quantized variant MiMo-VL-Miloco-7B-GGUF, a pair of home-centric vision-language models that achieve strong performance on both home-scenar…