collaborators

15 papers

cs.HC2026

HoosierHelp: Benchmarking LLM Agents for Social Service Navigation

Yiyang Li, Weixiang Sun, Tianyi Ma +3

Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM agents offer a promising interfa…

cs.LG2026

Policy4OOD: A Knowledge-Guided World Model for Policy Intervention Simulation against the Opioid Overdose Crisis

Yijun Ma, Zehong Wang, Weixiang Sun +4

The opioid epidemic remains one of the most severe public health crises in the United States, yet evaluating policy interventions before implementation is difficult: multiple polic…

cs.AI2026

Confidence Laundering in Agent Systems: Why Uncertainty Needs a Latent Carrier

Kaiwen Shi, Zheyuan Zhang, Han Bao +2

Modern agent systems can turn uncertainty into overconfidence. Fragile upstream decisions are often exposed to downstream components as clean intermediate artifacts, while the unce…

cs.CL2026

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

Kaiwen Shi, Zheyuan Zhang, Yanfang Ye

Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's sampled behavior. We study verba…

cs.CL2026

Food4All: An Agentic Framework and Benchmark for Food Resource Navigation with Adaptive User Understanding

Yiyang Li, Weixiang Sun, Tianyi Ma +3

Food assistance referral requires conversational agents to translate underspecified, often noisy help-seeking dialogues into locally valid resource recommendations. We present Food…

cs.LG2026

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

Zheyuan Zhang, Kaiwen Shi, Han Bao +3

Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from model-generated outputs but l…