1 citations · 1 across the 11 of their papers we have counts for
1 paper · 1 filter
Zhichao Fan, Yanhang Li, Zexin Zhuang
Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe evaluation recipe to test t…