Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Improving Vision-language Models with Perception-centric Process Reward Models
Yingqian Min, Kun Zhou, Yifan Li +6
Recent advancements in reinforcement learning with verifiable rewards (RLVR) have significantly improved the complex reasoning ability of vision-language models (VLMs). However, it…
cs.CV2026
SiMO: Single-Modality-Operable Multimodal Collaborative Perception
Jiageng Wen, Shengjie Zhao, Bing Li +3
Collaborative perception integrates multi-agent perspectives to enhance the sensing range and overcome occlusion issues. While existing multimodal approaches leverage complementary…