3 papers
cs.CV2025
MedGEN-Bench: A Contextually Entangled Benchmark for Open-ended Multimodal Medical Generation
Junjie Yang, Yuhao Yan, Gang Wu +12
Medical vision-language models (VLMs) are increasingly expected to support clinical workflows through diagnostic text and relevant medical images. However, current medical visual b…
cs.HC2024
Intuitive interaction flow: A Dual-Loop Human-Machine Collaboration Task Allocation Model and an experimental study
Jiang Xu, Qiyang Miao, Ziyuan Huang +5
This study investigates the issue of task allocation in Human-Machine Collaboration (HMC) within the context of Industry 4.0. By integrating philosophical insights and cognitive sc…
cs.CV2024
MambaBEV: An EV-based 3D detection model with Mamba2
Zihan You, Ni Wang, Hao Wang +2
Accurate 3D object detection in autonomous driving relies on Bird's Eye View (BEV) perception and effective temporal fusion. However, existing fusion strategies based on convolutio…