From the 1 of 1 linked paper with an AI index.
1 paper
Hiroto Osaka, Shohei Taniguchi, Gouki Minegishi +3
The paper investigates how chain-of-thought prompting works in vision-language models by introducing a visual access sweep that masks attention to image tokens, defining a visual a…