2 papers
cs.CV2026
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
Gracjan Góral, Alicja Ziarko, Piotr MiÅoÅ +3
We investigate the ability of Vision Language Models (VLMs) to perform visual perspective taking using a new set of visual tasks inspired by established human tests. Our approach l…
cs.SE2025
Measuring Determinism in Large Language Models for Software Code Review
Eugene Klishevich, Yegor Denisov-Blanch, Simon Obstbaum +2
Large Language Models (LLMs) promise to streamline software code reviews, but their ability to produce consistent assessments remains an open question. In this study, we tested fou…