2 papers
cs.LG2026
MemCalib: Benchmarking and Optimizing Memory Use in LLM Agents
Ruike Cao, Fanyu Zhao, Fugen Yao +6
The effectiveness of agent memory ultimately depends on whether the underlying LLM gives each memory in context an appropriate degree of influence over its response. Yet this capab…
cs.LG2026
ATPO: Adaptive Tree Policy Optimization for Multi-Turn Medical Dialogue
Ruike Cao, Shaojie Bai, Fugen Yao +3
Effective information seeking in multi-turn medical dialogues is critical for accurate diagnosis, especially when dealing with incomplete information. Aligning Large Language Model…