1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 1 cited
Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models
Zuyan Liu, Yuhao Dong, Yongming Rao +2
In the realm of vision-language understanding, the proficiency of models in interpreting and reasoning over visual content has become a cornerstone for numerous applications. Howev…
cs.CV2023★ 1 cited
NSM4D: Neural Scene Model Based Online 4D Point Cloud Sequence Understanding
Yuhao Dong, Zhuoyang Zhang, Yunze Liu +1
Understanding 4D point cloud sequences online is of significant practical value in various scenarios such as VR/AR, robotics, and autonomous driving. The key goal is to continuousl…