14 papers
HoosierHelp: Benchmarking LLM Agents for Social Service Navigation
Yiyang Li, Weixiang Sun, Tianyi Ma +3
Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM agents offer a promising interfa…
Object-Centric Environment Modeling for Agentic Tasks
Yiyang Li, Tianyi Ma, Zehong Wang +2
Large language model (LLM) agents can improve through accumulated experience, but free-form textual memories become difficult to maintain, validate, and reuse as interactions grow.…
ProPlay: Procedural World Models for Self-Evolving LLM Agents
Yijun Ma, Zehong Wang, Yiyang Li +5
Self-evolving agents are expected to improve through interaction without external supervision, but this remains difficult in partially observable environments where agents must exp…
Food4All: An Agentic Framework and Benchmark for Food Resource Navigation with Adaptive User Understanding
Yiyang Li, Weixiang Sun, Tianyi Ma +3
Food assistance referral requires conversational agents to translate underspecified, often noisy help-seeking dialogues into locally valid resource recommendations. We present Food…
PreScam: A Benchmark for Predicting Scam Progression from Early Conversations
Weixiang Sun, Shang Ma, Yiyang Li +5
Conversational scams, such as romance and investment scams, are emerging as a major form of online fraud. Unlike one-shot scam lures such as fake lottery or unpaid toll messages, t…
EvoTaxo: Building and Evolving Taxonomy from Social Media Streams
Yiyang Li, Tianyi Ma, Yanfang Ye
Constructing taxonomies from social media corpora is challenging because posts are short, noisy, semantically entangled, and temporally dynamic. Existing taxonomy induction methods…