From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
Consensus as Privileged Context for Label-Free Self-Distillation
John Gkountouras, Josip JukiÄ, Ivan Titov
The paper introduces CANON, a label‑free self‑distillation method that uses the majority answer from multiple sampled solutions as dense token‑level supervision to improve reasonin…
cs.CL2025
Controlling What You Share: Assessing Language Model Adherence to Privacy Preferences
Guillem RamÃrez, Alexandra Birch, Ivan Titov
Large language models (LLMs) are primarily accessed via commercial APIs, but this often requires users to expose their data to service providers. In this paper, we explore how user…
cs.LG2025
Clarification as Supervision: Reinforcement Learning for Vision-Language Interfaces
John Gkountouras, Ivan Titov
Recent text-only models demonstrate remarkable mathematical reasoning capabilities. Extending these to visual domains requires vision-language models to translate images into text…