2 papers
cs.CL2024
Large Language Model Instruction Following: A Survey of Progresses and Challenges
Renze Lou, Kai Zhang, Wenpeng Yin
Task semantics can be expressed by a set of input-output examples or a piece of textual instruction. Conventional machine learning approaches for natural language processing (NLP)…
cs.CL2024
Anti-Overestimation Dialogue Policy Learning for Task-Completion Dialogue System
Chang Tian, Wenpeng Yin, Marie-Francine Moens
A dialogue policy module is an essential part of task-completion dialogue systems. Recently, increasing interest has focused on reinforcement learning (RL)-based dialogue policy. I…