From the 1 of 7 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Consensus as Privileged Context for Label-Free Self-Distillation
John Gkountouras, Josip JukiÄ, Ivan Titov
The paper introduces CANON, a label‑free self‑distillation method that uses the majority answer from multiple sampled solutions as dense token‑level supervision to improve reasonin…
cs.LG2025
Clarification as Supervision: Reinforcement Learning for Vision-Language Interfaces
John Gkountouras, Ivan Titov
Recent text-only models demonstrate remarkable mathematical reasoning capabilities. Extending these to visual domains requires vision-language models to translate images into text…