1 citations · 1 across the 1 of their papers we have counts for
3 papers
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
Gracjan Góral, Alicja Ziarko, Piotr Miłoś +3
We investigate the ability of Vision Language Models (VLMs) to perform visual perspective taking using a new set of visual tasks inspired by established human tests. Our approach l…
Measuring Determinism in Large Language Models for Software Code Review
Eugene Klishevich, Yegor Denisov-Blanch, Simon Obstbaum +2
Large Language Models (LLMs) promise to streamline software code reviews, but their ability to produce consistent assessments remains an open question. In this study, we tested fou…
Predicting Expert Evaluations in Software Code Reviews
Yegor Denisov-Blanch, Igor Ciobanu, Simon Obstbaum +1
Manual code reviews are an essential but time-consuming part of software development, often leading reviewers to prioritize technical issues while skipping valuable assessments. Th…