From the 1 of 10 linked papers with an AI index.
10 papers
LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning
Ziyan Xiao, Yinghao Zhu, Wenting Zhang +2
Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart understanding and multi-step i…
Evidence-Grounded AI for Musculoskeletal Care
Wenjie Li, Yujie Zhang, Fanrui Zhang +34
The paper presents OrthoPilot, a clinical AI system powered by a large language model that integrates real-time hospital data and external medical knowledge to provide evidence‑bas…
Auditing medical multi-agent AI reveals risks of false consensus
Yinghao Zhu, Lei Gu, Zixiang Wang +11
Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through specialist roles, peer review and consen…
CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation
Yushi Feng, Junye Du, Qifan Wang +5
Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. However, this unrestricted action…
Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage
Ziyi He, Yushi Feng, Shuangyu Yang +7
Dental triage is a safety-critical clinical routing task that requires integrating multimodal clinical information (e.g., patient complaints and radiographic evidence) to determine…
MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing
Ziyan Xiao, Yinghao Zhu, Liang Peng +2
Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this task remains challenging because…