4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2026★ 4 cited
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
Haoxuan You, Rui Sun, Zhecan Wang +5
The field of vision-and-language (VL) understanding has made unprecedented progress with end-to-end large pre-trained VL models (VLMs). However, they still fall short in zero-shot…
cs.IR2026
Libra: Training the Environment for Agentic Information Retrieval
Xuan Zhao, Andy Chiu, Gengyu Wang
Information localization within massive repositories is a cornerstone of agentic LLM systems. While synthetic data-driven optimization has proven successful in training LLMs, littl…