Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Plan Before Search: Search Agents Need Plan
Zhipeng Qian, Zihan Liang, Yufei Ma +7
Training large language models as retrieval-augmented reasoning agents typically combines reinforcement learning with an SFT cold start distilled from a stronger model. However, th…
cs.AI2026
GeoAgent: Learning to Geolocate Everywhere with Reinforced Geographic Characteristics
Modi Jin, Yiming Zhang, Boyuan Sun +3
This paper presents GeoAgent, a model capable of reasoning closely with humans and deriving fine-grained address conclusions. Previous RL-based methods have achieved breakthroughs…