30 citations · 30 across the 2 of their papers we have counts for
2 papers
cs.CV2026
VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning
Shufan Zhang, Ziyue Lin, Bairun Wang +4
Video reasoning aims to understand complex temporal events and causal relationships within videos. Recently, Chain-of-Thought (CoT) has been introduced to this field to enhance rea…
cs.CV2020★ 30 cited
BriNet: Towards Bridging the Intra-class and Inter-class Gaps in One-Shot Segmentation
Xianghui Yang, Bairun Wang, Kaige Chen +4
Few-shot segmentation focuses on the generalization of models to segment unseen object instances with limited training samples. Although tremendous improvements have been achieved,…