10 citations · 23 across the 10 of their papers we have counts for
12 papers · 1 filter
Distributional Clarity: The Hidden Driver of RL-Friendliness in Large Language Models
Shaoning Sun, Mingzhu Cai, Huang He +5
Language model families exhibit striking disparity in their capacity to benefit from reinforcement learning: under identical training, models like Qwen achieve substantial gains, w…
ProxyAttn: Guided Sparse Attention via Representative Heads
Yixuan Wang, Huang He, Siqi Bao +4
The quadratic complexity of attention mechanisms limits the efficiency of Large Language Models (LLMs) on long-text tasks. Recently, methods that dynamically estimate block importa…
PLATO-K: Internal and External Knowledge Enhanced Dialogue Generation
Siqi Bao, Huang He, Jun Xu +7
Recently, the practical deployment of open-domain dialogue systems has been plagued by the knowledge issue of information deficiency and factual inaccuracy. To this end, we introdu…
Q-TOD: A Query-driven Task-oriented Dialogue System
Xin Tian, Yingzhan Lin, Mengfei Song +5
Existing pipelined task-oriented dialogue systems usually have difficulties adapting to unseen domains, whereas end-to-end systems are plagued by large-scale knowledge bases in pra…
Towards Building an Open-Domain Dialogue System Incorporated with Internet Memes
Hua Lu, Zhen Guo, Chanjuan Li +3
In recent years, Internet memes have been widely used in online chatting. Compared with text-based communication, conversations become more expressive and attractive when Internet…
Amendable Generation for Dialogue State Tracking
Xin Tian, Liankai Huang, Yingzhan Lin +6
In task-oriented dialogue systems, recent dialogue state tracking methods tend to perform one-pass generation of the dialogue state based on the previous dialogue state. The mistak…