6 papers
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
Gracjan Góral, Alicja Ziarko, Piotr MiÅoÅ +3
We investigate the ability of Vision Language Models (VLMs) to perform visual perspective taking using a new set of visual tasks inspired by established human tests. Our approach l…
OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
PaweÅ Budzianowski, Emilia WiÅnios, MichaÅ Tyrolski +4
Data scarcity remains one of the most limiting factors in driving progress in robotics. However, the amount of available robotics data in the wild is growing exponentially, creatin…
Depth-Wise Activation Steering for Honest Language Models
Gracjan Góral, Marysia Winkels, Steven Basart
Large language models sometimes assert falsehoods despite internally representing the correct answer, failures of honesty rather than accuracy, which undermines auditability and sa…
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
Gracjan Góral, Emilia WiÅnios, Piotr Sankowski +1
This work introduces a novel framework for evaluating LLMs' capacity to balance instruction-following with critical reasoning when presented with multiple-choice questions containi…
What Matters in Hierarchical Search for Combinatorial Reasoning Problems?
MichaŠZawalski, Gracjan Góral, MichaŠTyrolski +5
Efficiently tackling combinatorial reasoning problems, particularly the notorious NP-hard tasks, remains a significant challenge for AI research. Recent efforts have sought to enha…
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
Gracjan Góral, Alicja Ziarko, Michal Nauman +1
Visual perspective-taking (VPT), the ability to understand the viewpoint of another person, enables individuals to anticipate the actions of other people. For instance, a driver ca…