2 papers
cs.CL2026
DPlan: Dual-Agent Dynamic Global Planning for Complex Retrieval-Augmented Reasoning
Kangcheng Luo, Tinglang Wu, Yansong Feng
Recent search-augmented LLMs trained with reinforcement learning (RL) can interleave searching and reasoning for multi-hop reasoning tasks. However, they face two critical failure…
cs.CL2025
BEDA: Belief Estimation as Probabilistic Constraints for Performing Strategic Dialogue Acts
Hengli Li, Zhaoxin Yu, Qi Shen +8
Strategic dialogue requires agents to execute distinct dialogue acts, for which belief estimation is essential. While prior work often estimates beliefs accurately, it lacks a prin…