Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents
Zhixiang Liang, Yifei Liu, Yidan Huang +5
Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning errors may propagate through long…
cs.AI2026
SearchMaster: Grounded and Regulated Self-Play for Search Agents
Wentao Tan, Qiong Cao, Jiaqi Wang +1
Training LLM-based search agents requires high-quality search data: tasks that demand genuine multi-hop retrieval and trajectories that use search tools effectively. Existing pipel…
cs.AI2025
Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs
Chang Li, Yaren Zhang, Haoran Lv +3
Large Language Models (LLMs) have shown remarkable reasoning ability through explicit Chain-of-Thought (CoT) prompting, but generating these step-by-step textual explanations is co…