4 citations · 4 across the 9 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.CV2025
CoRe3D: Collaborative Reasoning as a Foundation for 3D Intelligence
Tianjiao Yu, Xinzhuo Li, Yifan Shen +2
Recent advances in large multimodal models suggest that explicit reasoning mechanisms play a critical role in improving model reliability, interpretability, and cross-modal alignme…
cs.CV2025
Fine-Grained Preference Optimization Improves Spatial Reasoning in VLMs
Yifan Shen, Yuanzhe Liu, Jingyuan Zhu +6
Current Vision-Language Models (VLMs) struggle with fine-grained spatial reasoning, particularly when multi-step logic and precise spatial alignment are required. In this work, we…