collaborators

22 papers

cs.LG2026

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

Zheyuan Zhang, Manqing Mao, Hong Wang +8

Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods allocate the same number of ro…

cs.HC2026

HoosierHelp: Benchmarking LLM Agents for Social Service Navigation

Yiyang Li, Weixiang Sun, Tianyi Ma +3

Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM agents offer a promising interfa…

cs.AI2026

Object-Centric Environment Modeling for Agentic Tasks

Yiyang Li, Tianyi Ma, Zehong Wang +2

Large language model (LLM) agents can improve through accumulated experience, but free-form textual memories become difficult to maintain, validate, and reuse as interactions grow.…

cs.LG2026

Policy4OOD: A Knowledge-Guided World Model for Policy Intervention Simulation against the Opioid Overdose Crisis

Yijun Ma, Zehong Wang, Weixiang Sun +4

The opioid epidemic remains one of the most severe public health crises in the United States, yet evaluating policy interventions before implementation is difficult: multiple polic…

cs.LG2026

Temporal Graph Pattern Machine

Yijun Ma, Zehong Wang, Weixiang Sun +1

Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving patterns that govern network transf…

cs.LG2026

SupraBench: A Benchmark for Supramolecular Chemistry

Tianyi Ma, Yijun Ma, Zehong Wang +6

Supramolecular chemistry, which includes the study of non-covalent host-guest assemblies, has advanced various applications. However, designing host-guest systems remains time-cons…