agent architecture 1longitudinal memory 1medical knowledge integration 1personal health management 1privacy-aware AI 1
From the 1 of 7 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Evaluating Stochastic Collapse and Implicit Bias in Multimodal Large Language Models
Huiyuan Zheng, Houtao Zhang, Boyang Wang +2
Current evaluations for Multimodal Large Language Models (MLLMs) overwhelmingly focus on utility-driven objectives, leaving model behavior under logic-neutral scenarios largely und…
cs.CL2026
Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning
Ziheng Li, Liu Kang, Feng Xiao +7
Group Relative Policy Optimization (GRPO) has emerged as a promising critic-free reinforcement learning paradigm for reasoning tasks. However, standard GRPO employs a coarse-graine…