2 papers
cs.CL2025
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Kaiwen Zuo, Yirui Jiang
Medical Large Language Models (MLLMs) have demonstrated potential in healthcare applications, yet their propensity for hallucinations -- generating medically implausible or inaccur…
cs.AI2025
KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Medical Diagnosis
Kaiwen Zuo, Yirui Jiang, Fan Mo +1
Integrating Large Language Models (LLMs) in healthcare diagnosis demands systematic frameworks that can handle complex medical scenarios while maintaining specialized expertise. We…