2 papers
cs.CL2025
Why Chain of Thought Fails in Clinical Text Understanding
Jiageng Wu, Kevin Xie, Bowen Gu +3
Large language models (LLMs) are increasingly being applied to clinical care, a domain where both accuracy and transparent reasoning are critical for safe and trustworthy deploymen…
cs.CL2025
BRIDGE: Benchmarking Large Language Models for Understanding Real-world Clinical Practice Text
Jiageng Wu, Bowen Gu, Ren Zhou +14
Large language models (LLMs) hold great promise for medical applications and are evolving rapidly, with new models being released at an accelerated pace. However, benchmarking on l…