3 papers
cs.CV2026
Local Margin Restoration for Test-Time Adaptation of Vision-Language Models
Yan Huang, Guowei Wang, Xu Wang +2
Vision-language models (VLMs) such as CLIP exhibit remarkable zero-shot capabilities, yet their performance frequently degrades sharply under unexpected test-time distribution shif…
cs.CL2026
LEEPS: Latent-Guided Explore-Exploit Prompt Sampling for Efficient RLVR in Large Language Models
Shuang Liang, Haoyang Zhou, Yifan Gong +2
Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models, but prompt groups with identical rollout rewards consume generat…
cs.CV2025
Effortless Active Labeling for Long-Term Test-Time Adaptation
Guowei Wang, Changxing Ding
Long-term test-time adaptation (TTA) is a challenging task due to error accumulation. Recent approaches tackle this issue by actively labeling a small proportion of samples in each…