4 papers · 1 filter
MentalHospital: A Virtual Environment for Evaluating Psychiatric Clinical Encounters
Yuming Yang, Xiao Sun, Yuanwei Zou +6
Large language models (LLMs) have shown strong performance on isolated psychiatric tasks, including dialogue, diagnosis, and treatment planning, yet existing benchmarks rarely simu…
Lung-R1: A Knowledge Graph-Guided LLM for Pulmonary Diagnostic Reasoning
Haoyang Zeng, Yuanxi Fu, Rongzhen Li +11
Diagnosing pulmonary diseases requires integrating heterogeneous evidence amid phenotypic variability and cross-disease overlap. Although large language models (LLMs) have shown pr…
MentalSeek-Dx: Towards Progressive Hypothetico-Deductive Reasoning for Real-world Psychiatric Diagnosis
Xiao Sun, Yuming Yang, Junnan Zhu +3
Mental health disorders represent a burgeoning global public health challenge. While Large Language Models (LLMs) have demonstrated potential in psychiatric assessment, their clini…
Benchmarking Multimodal RAG through a Chart-based Document Question-Answering Generation Framework
Yuming Yang, Jiang Zhong, Li Jin +7
Multimodal Retrieval-Augmented Generation (MRAG) enhances reasoning capabilities by integrating external knowledge. However, existing benchmarks primarily focus on simple image-tex…