11 papers
CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation
Yaoning Yu, Kai-Min Chang, Ye Yu +3
Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions requires more than just gener…
QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing
Yijing Zuo, Zhe Fu, Zihan Nie +2
Identifying a meaningful quantum speedup requires more than matching a classical problem to a familiar quantum primitive: the claim must preserve the task, respect access and outpu…
KV-PRM: Efficient Process Reward Modeling via KV-Cache Transfer for Multi-Agent Test-Time Scaling
Peng Kuang, Haibo Jin, Xiaoyu Han +5
Process Reward Models (PRMs) have been proven to be highly effective in guiding test-time scaling (TTS) methods, which significantly boost the capabilities of LLM-based multi-agent…
Closing the Loop on Latent Reasoning via Test-Time Reconstruction
Xiaopeng Yuan, Haibo Jin, Ye Yu +4
Recent work moves intermediate reasoning from natural-language traces into latent or cache-level representations to reduce token overhead and avoid a discrete communication bottlen…
MiroBench: Benchmarking Realism in Agentic Simulation of Real-world Discussions
Yaoning Yu, Ye Yu, Haojing Luo +1
LLM agents are increasingly used to simulate real world interactions, but it remains unclear whether simulated behaviors preserve the content patterns and interaction dynamics of r…
Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation
Ye Yu, Xiaopeng Yuan, Haibo Jin +3
Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and maintain persistent memory. How…