2 papers
cs.CV2026
GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning
Brown Ebouky, Gabriele Carrino, Niccolo Avogaro +3
Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routing foveal focus toward task-rel…
cs.LG2026
Are complicated loss functions necessary for teaching LLMs to reason?
Gabriele Carrino, Andrea Sassella, Nicolo Brunello +2
Recent advances in large language models (LLMs) highlight the importance of post training techniques for improving reasoning and mathematical ability. Group Relative Policy Optimiz…