From the 1 of 25 linked papers with an AI index.
4 papers · 1 filter
MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning
Haotian Wang, Lian Yan, Xingzhi Yao +4
In Reinforcement Learning with Verifiable Rewards (RLVR) frameworks for mathematical reasoning tasks, floating-point results are typically evaluated using a tolerance-based reward.…
RealUnify: Do Unified Models Truly Benefit from Unification? A Comprehensive Benchmark
Yang Shi, Yuhao Dong, Yue Ding +22
The integration of visual understanding and generation into unified multimodal models represents a significant stride toward general-purpose AI. However, a fundamental question rem…
Coordinated Pandemic Control with Large Language Model Agents as Policymaking Assistants
Ziyi Shi, Xusen Guo, Hongliang Lu +7
Effective pandemic control requires timely and coordinated policymaking across administrative regions that are intrinsically interdependent. However, human-driven responses are oft…
When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs
Zhuoran Zhang, Tengyue Wang, Xilin Gong +4
Multimodal large language models (MLLMs) must resolve conflicts when different modalities provide contradictory information, a process we term modality following. Prior work measur…