1 paper · 1 filter
Jonathan Cook, Diego Antognini, Martin Klissarov +2
Large language models (LLMs) often struggle to learn from corrective feedback within a conversational context. They are rarely proactive in soliciting this feedback, even when face…