Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Optimizing Agentic Reasoning with Retrieval via Synthetic Semantic Information Gain Reward
Senkang Hu, Yong Dai, Yuzhi Zhao +5
Agentic reasoning enables large reasoning models (LRMs) to dynamically acquire external knowledge, but yet optimizing the retrieval process remains challenging due to the lack of d…
cs.AI2026
Inference-Time Budget Control for LLM Search Agents
Zhengru Fang, Senkang Forest Hu, Zhonghao Chang +6
LLM search agents increasingly rely on tools at inference time, but their trajectories are often constrained by hard limits on both tool calls and generated tokens. Under such dual…
cs.AI2026
Shared Spatial Memory Through Predictive Coding
Zhengru Fang, Yu Guo, Yuang Zhang +3
Constructing a consistent shared spatial memory is a critical challenge in multi-agent systems, where partial observability and limited bandwidth often lead to catastrophic failure…