2 papers
cs.LG2025
Enhancing RLHF with Human Gaze Modeling
Karim Galliamov, Ivan Titov, Ilya Pershin
Reinforcement Learning from Human Feedback (RLHF) aligns language models with human preferences but is computationally expensive. We explore two approaches that leverage human gaze…
cs.CL2023
Cross-Modal Conceptualization in Bottleneck Models
Danis Alukaev, Semen Kiselev, Ilya Pershin +4
Concept Bottleneck Models (CBMs) assume that training examples (e.g., x-ray images) are annotated with high-level concepts (e.g., types of abnormalities), and perform classificatio…