Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning
Chaofan Ma, Zhenjie Mao, Yuhuan Yang +5
Spatial reasoning from egocentric videos is inherently challenging because the observable evidence is constrained by the camera trajectory. Existing methods rely on single-turn inf…
cs.CV2025
An Empirical Analysis of VLM-based OOD Detection: Mechanisms, Advantages, and Sensitivity
Yuxiao Lee, Xiaofeng Cao, Wei Ye +3
Vision-Language Models (VLMs), such as CLIP, have demonstrated remarkable zero-shot out-of-distribution (OOD) detection capabilities, vital for reliable AI systems. Despite this pr…