From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
Consensus as Privileged Context for Label-Free Self-Distillation
John Gkountouras, Josip JukiÄ, Ivan Titov
The paper introduces CANON, a label‑free self‑distillation method that uses the majority answer from multiple sampled solutions as dense token‑level supervision to improve reasonin…
cs.LG2025
Clarification as Supervision: Reinforcement Learning for Vision-Language Interfaces
John Gkountouras, Ivan Titov
Recent text-only models demonstrate remarkable mathematical reasoning capabilities. Extending these to visual domains requires vision-language models to translate images into text…
cs.AI2024
Language Agents Meet Causality -- Bridging LLMs and Causal World Models
John Gkountouras, Matthias Lindemann, Phillip Lippe +2
Large Language Models (LLMs) have recently shown great promise in planning and reasoning applications. These tasks demand robust systems, which arguably require a causal understand…