1 paper
Jinhwan Seo, Kyubeom Han, Jumin Lee +2
We study a critical yet overlooked failure mode in Grounded Video Question Answering: question-invariant grounding, where models predict nearly identical temporal segments for diff…