110 citations · 114 across the 15 of their papers we have counts for
11 papers · 1 filter
ToSCA: Leveraging Hierarchical Reinforcement Learning on Temporal and Strategic Abstractions of Conversational Agents
Xiaoyu Wang, Qingqing Gu, Yue Zhao +5
Humans naturally exhibit multiple forms of abstraction in reasoning and interaction, including temporal abstraction across decision timescales and strategic abstraction over commun…
Learn-To-Learn on Arbitrary Textual Conditioning: A Hypernetwork-Driven Meta-Gated LLM
Luo Ji, Qi Qin, Ningyuan Xi +3
Conventional LLMs may suffer from corpus heterogeneity and subtle condition changes. While finetuning can create the catastrophe forgetting issue, application of meta-learning on L…
Chain-of-Conceptual-Thought Elicits Daily Conversation in Large Language Models
Qingqing Gu, Dan Wang, Yue Zhao +5
Chain-of-Thought (CoT) is widely applied to enhance the LLM capability in math, coding and reasoning tasks. However, its performance is limited for open-domain tasks, when there ar…
Dream to Chat: Model-based Reinforcement Learning on Dialogues with User Belief Modeling
Yue Zhao, Xiaoyu Wang, Dan Wang +7
World models have been widely utilized in robotics, gaming, and auto-driving. However, their applications on natural language tasks are relatively limited. In this paper, we constr…
Convert Language Model into a Value-based Strategic Planner
Xiaoyu Wang, Yue Zhao, Qingqing Gu +4
Emotional support conversation (ESC) aims to alleviate the emotional distress of individuals through effective conversations. Although large language models (LLMs) have obtained re…
EmoFSM: A Finite State Machine for Emotional Support Conversation
Yue Zhao, Qingqing Gu, Xiaoyu Wang +5
Emotional support conversation (ESC) aims to alleviate people's emotional distress through effective conversations. Although large language models (LLMs) have made remarkable progr…