From the 1 of 4 linked papers with an AI index.
7 papers · 1 filter
REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering
Li-Ming Zhan, Bo Liu, Chengqiang Xie +2
The paper introduces REAL, a method that trains vector-quantized autoencoders on transformer activations to pinpoint attention heads or layers that most influence a target behavior…
KIF: Knowledge Identification and Fusion for Language Model Continual Learning
Yujie Feng, Xu Chu, Yongxin Xu +4
Language model continual learning (CL) has recently attracted significant interest for its ability to adapt large language models (LLMs) to dynamic real-world scenarios without ret…
Continual Dialogue State Tracking via Reason-of-Select Distillation
Yujie Feng, Bo Liu, Xiaoyu Dong +4
An ideal dialogue system requires continuous skill acquisition and adaptation to new tasks while retaining prior knowledge. Dialogue State Tracking (DST), vital in these systems, o…
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
Bo Liu, Liming Zhan, Yujie Feng +5
In the realm of task-oriented dialogue systems, a robust intent detection mechanism must effectively handle malformed utterances encountered in real-world scenarios. This study pre…
TaSL: Continual Dialog State Tracking via Task Skill Localization and Consolidation
Yujie Feng, Xu Chu, Yongxin Xu +3
A practical dialogue system requires the capacity for ongoing skill acquisition and adaptability to new tasks while preserving prior knowledge. However, current methods for Continu…
How Good Are LLMs at Out-of-Distribution Detection?
Bo Liu, Liming Zhan, Zexin Lu +3
Out-of-distribution (OOD) detection plays a vital role in enhancing the reliability of machine learning (ML) models. The emergence of large language models (LLMs) has catalyzed a p…