2 papers
cs.CV2026
Training-Free VLM Personalization via Calibrated Residual Decoding
Jiaao Yu, Yujian Ma, Xianming Hu +2
Vision-language models can be personalized in a training-free manner by directly providing user profiles, preferences, or visual references at inference time, without updating mode…
cs.CL2026
Reward Prediction with Factorized World States
Yijun Shen, Delong Chen, Xianming Hu +4
Agents must infer action outcomes and select actions that maximize a reward signal indicating how close the goal is to being reached. Supervised learning of reward models could int…