31 papers
Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning
Zheyuan Zhang, Manqing Mao, Hong Wang +8
Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods allocate the same number of ro…
HoosierHelp: Benchmarking LLM Agents for Social Service Navigation
Yiyang Li, Weixiang Sun, Tianyi Ma +3
Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM agents offer a promising interfa…
Policy4OOD: A Knowledge-Guided World Model for Policy Intervention Simulation against the Opioid Overdose Crisis
Yijun Ma, Zehong Wang, Weixiang Sun +4
The opioid epidemic remains one of the most severe public health crises in the United States, yet evaluating policy interventions before implementation is difficult: multiple polic…
Confidence Laundering in Agent Systems: Why Uncertainty Needs a Latent Carrier
Kaiwen Shi, Zheyuan Zhang, Han Bao +2
Modern agent systems can turn uncertainty into overconfidence. Fragile upstream decisions are often exposed to downstream components as clean intermediate artifacts, while the unce…
SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment
Kaiwen Shi, Zheyuan Zhang, Yanfang Ye
Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's sampled behavior. We study verba…
Food4All: An Agentic Framework and Benchmark for Food Resource Navigation with Adaptive User Understanding
Yiyang Li, Weixiang Sun, Tianyi Ma +3
Food assistance referral requires conversational agents to translate underspecified, often noisy help-seeking dialogues into locally valid resource recommendations. We present Food…