2 papers
cs.CL2026
Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach
Tianbao Jiang, Weicong Ni, Gerard de Melo +1
Post-training reinforcement learning (RL) algorithms are commonly used to align large vision-language models (LVLMs) with human intent and the requirements of visual reasoning task…
cs.AI2026
Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models
Weicong Ni, Tianbao Jiang, Linlin Wang
Vision-Language Models (VLMs) are becoming the cornerstone of high-level reasoning for robotic automation, enabling robots to parse natural language commands and perceive their env…