1 citations · 1 across the 10 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
CDEG: Learning Decision-Critical Evidence for Long-Horizon Diagnostic Agents
Xiwei Dai, Zijie Meng, Zhiting Fan +4
Unlike static medical question answering, long-horizon diagnosis captures the sequential nature of clinical practice: evidence is progressively acquired, integrated, and evaluated…
cs.AI2026
LongMedBench: Benchmarking Medical Agents for Long-Horizon Clinical Decision-Making
Zihan Xu, Yanzhen Chen, Xiaocheng Zhang +4
In this work, we introduce LongMedBench, a real-world EHR-based benchmark for long-horizon clinical decision-making. Prior evaluations of LLM-based medical agents have largely emph…
cs.AI2025★ 1 cited
Reinforcement Learning with Rubric Anchors
Zenan Huang, Yihong Zhuang, Guoshan Lu +18
Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for enhancing Large Language Models (LLMs), exemplified by the success of OpenAI's o-series…