Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
ToSCA: Leveraging Hierarchical Reinforcement Learning on Temporal and Strategic Abstractions of Conversational Agents
Xiaoyu Wang, Qingqing Gu, Yue Zhao +5
Humans naturally exhibit multiple forms of abstraction in reasoning and interaction, including temporal abstraction across decision timescales and strategic abstraction over commun…
cs.CL2025
Chain-of-Conceptual-Thought Elicits Daily Conversation in Large Language Models
Qingqing Gu, Dan Wang, Yue Zhao +5
Chain-of-Thought (CoT) is widely applied to enhance the LLM capability in math, coding and reasoning tasks. However, its performance is limited for open-domain tasks, when there ar…