works on

From the 1 of 10 linked papers with an AI index.

collaborators

10 papers

cs.CL2026

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

Ziyan Xiao, Yinghao Zhu, Wenting Zhang +2

Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart understanding and multi-step i…

cs.AI2026

Evidence-Grounded AI for Musculoskeletal Care

Wenjie Li, Yujie Zhang, Fanrui Zhang +34

The paper presents OrthoPilot, a clinical AI system powered by a large language model that integrates real-time hospital data and external medical knowledge to provide evidence‑bas…

cs.CL2026

Auditing medical multi-agent AI reveals risks of false consensus

Yinghao Zhu, Lei Gu, Zixiang Wang +11

Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through specialist roles, peer review and consen…

cs.LG2026

CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation

Yushi Feng, Junye Du, Qifan Wang +5

Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. However, this unrestricted action…

cs.CL2026

Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage

Ziyi He, Yushi Feng, Shuangyu Yang +7

Dental triage is a safety-critical clinical routing task that requires integrating multimodal clinical information (e.g., patient complaints and radiographic evidence) to determine…

cs.CL2026

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

Ziyan Xiao, Yinghao Zhu, Liang Peng +2

Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this task remains challenging because…