14 papers
SocialBuddy: Tailoring Search Agent for Social Scenarios
Mingxuan Li, Yirong Mao, FaZhan Zhang +3
In the era of digital social interaction, searching friends' posts from massive social streams has become a fundamental user need. However, while modern agentic search frameworks h…
IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents
Bo Ren, Yirong Mao, Yi Yang +1
Large Language Model (LLM) agents increasingly solve long-horizon tasks through multi-turn interactions with users and external tools. In these settings, relevant task information…
EviGraph: Towards Verifiable Evidence Construction for Information-Seeking Agents
Jiashun Chen, Yirong Mao, Wenhui Que
Agentic Web search can retrieve relevant information without establishing that the retrieved content actually supports the claims used in an answer. Existing agents typically keep…
EvoWiki: Incremental State Overwriting and Traceable Question Answering for Cross-Meeting Knowledge Evolution
Dongsheng Chen, Tianyu Wang, Wenhui Que
In long-term collaboration spanning multiple meetings, factual states such as decisions and risks are continually revised, overturned, and replaced. Existing long-context methods t…
When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories
Xiaoqing Wu, Xingyu Fan, Feifei Li +1
Tool-calling agents infer task state from accumulated dialogue and tool traces. In persistent interactions, however, historical traces may remain structurally valid and semanticall…
DUET: Dual-Teacher On-Policy Distillation via Same-Weight Disagreement for Prohibition Compliance
Zihan Li, Feifei Li, Wenhui Que
Real-world LLM deployments increasingly rely on runtime-injected prohibitions--enterprise policies, PII redlines, tool boundaries--that vary per request and per tenant. Conventiona…