2 papers
cs.AI2026
APEx: Distillation of Agent Procedural Experience for Adaptive Deep Research Question Answering
Jie Ding, Rui Sun, Xinyuan Zhang +2
Deep research agents augment large language models with external tools to answer complex, long-horizon questions through multi-turn reasoning. Learning from prior experience is cru…
cs.CV2024
Tuning Vision-Language Models with Candidate Labels by Prompt Alignment
Zhifang Zhang, Yuwei Niu, Xin Liu +1
Vision-language models (VLMs) can learn high-quality representations from a large-scale training dataset of image-text pairs. Prompt learning is a popular approach to fine-tuning V…