2 papers
cs.CL2024
Aligning LLM Agents by Learning Latent Preference from User Edits
Ge Gao, Alexey Taymanov, Eduardo Salinas +2
We study interactive learning of LLM-based language agents based on user edits made to the agent's output. In a typical setting such as writing assistants, the user interacts with…
cs.LG2024
Policy Improvement using Language Feedback Models
Victor Zhong, Dipendra Misra, Xingdi Yuan +1
We introduce Language Feedback Models (LFMs) that identify desirable behaviour - actions that help achieve tasks specified in the instruction - for imitation learning in instructio…