1 paper
Yilin Ou, Mengshi Qi, Huadong Ma
The VRR-QA challenge evaluates visual relational reasoning in videos, where answers often depend on implicit spatial relations, event boundaries, target identity, and dialogue cont…