1 paper · 1 filter
Zhixiang Wang, Jingxuan Xu, Dajun Chen +3
Recent advances in Vision-Language Models (VLMs) have motivated the development of multi-modal search agents that can actively invoke external search tools and integrate retrieved…