3 papers
cs.CV2026
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding
Zhixuan Wu, Quanxing Zha, Teng Wang +6
Video understanding requires identifying and reasoning over semantically discriminative visual objects across frames, yet existing object-agnostic solutions struggle to effectively…
cs.CV2025
VRS-UIE: Value-Driven Reordering Scanning for Underwater Image Enhancement
Kui Jiang, Yan Luo, Junjun Jiang +3
State Space Models (SSMs) have emerged as a promising backbone for vision tasks due to their linear complexity and global receptive field. However, in the context of Underwater Ima…
cs.RO2024
AHPPEBot: Autonomous Robot for Tomato Harvesting based on Phenotyping and Pose Estimation
Xingxu Li, Nan Ma, Yiheng Han +2
To address the limitations inherent to conventional automated harvesting robots specifically their suboptimal success rates and risk of crop damage, we design a novel bot named AHP…